this post was submitted on 11 Sep 2026
89 points (83.0% liked)

Technology

87980 readers
2466 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
top 13 comments
sorted by: hot top controversial new old
[–] yardratianSoma@lemmy.ca 9 points 15 hours ago (1 children)

ummm, the only times I ever used deepseek was fully offline on my hardware, I call bullshit

[–] Womble@piefed.world 1 points 13 hours ago (1 children)

You almost certainly weren't using a Deepseek model offline (unless you have 100s of GB of RAM/VRAM). You might have been using a small model distilled on Deepseek responses, but the smallest model made by Deepseek is about a 300B param model.

[–] SteveTech@aussie.zone 7 points 12 hours ago

the smallest model made by Deepseek is about a 300B param model.

This is not correct. DeepSeek has plenty of smaller models. I just had a scroll through their huggingface, and picked out a few: DeepSeek-V2-Lite (16B), deepseek-moe-16b-chat, deepseek-llm-7b-chat, but there's quite a few more too.

Though DeepSeek R1 is their flagship model, which is the one people usually think of, and that's a 684B param model.

[–] FirmDistribution@lemmy.world 62 points 23 hours ago

is anthropic.com a reliable source for an information that benefits anthropic's image?

[–] pulsewidth@lemmy.world 36 points 23 hours ago (1 children)
[–] uuj8za@piefed.social 5 points 18 hours ago

Proof: Claude told me.

[–] Mustachius_Grumpius@thelemmy.club 58 points 1 day ago (1 children)

I don't believe a word they say.

They are lying for their stupid IPO

[–] Zetta@mander.xyz 4 points 23 hours ago

I love open Chinese llms but this seems 100% plausible and and I think ur wring. This is in line with how Chinese companies need to work around the system since America tries to handicap them. This is actually a very smart way to get high quality trianjng data for their future models. Unfortunate about leaking private customer data to a us company though.

Also the article and the evidence presented by Antropic all seems very plausible.

[–] civ@lemmy.civl.cc 79 points 1 day ago
[–] trolololol@lemmy.world 43 points 1 day ago

The onion is nothing compared to real headlines

[–] francisco_1844@discuss.online 8 points 21 hours ago

What this really boils down to is that Anthropic and OpenAI hope to be the only ones that use someone else's copyrighted information without any legal or financial complications.

[–] esc@piefed.social 8 points 23 hours ago* (last edited 23 hours ago)

Alright, so what? Also this (even if it's true and not another clown show by claude) won't affect models that aren't hosted by moonshot or deepseek.

[–] rimu@piefed.social -5 points 1 day ago