this post was submitted on 25 Jul 2026
631 points (98.3% liked)

Technology

86633 readers
3225 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] boonhet@sopuli.xyz 15 points 1 day ago (3 children)

Open-weight not open source mostly. And they're not keeping them open weight forever. Once they've truly surpassed the west, their future models will likely be closed down too so they could charge Anthropic level pricing and turn a profit.

[–] Guilvareux@feddit.uk 3 points 21 hours ago (1 children)

I don’t know how much they’re actually doing to surpass the West tbh. A lot of the success relies on distilling closed models, and running the models cheaper with almost comparable effectiveness. If the Chinese are showing that such profit/investment strategies in closed models are weak and can be easily decimated by a competitor, what motivation would they have to adopt one?

[–] boonhet@sopuli.xyz 4 points 20 hours ago (1 children)

If their models truly are reaching parity with new western models as many claim, then it can't be from distillation alone, they must be gathering their own datasets too. It takes months to train a new model.

Also distillation only really saves you the data collection and preparation (categorization). Training is still expensive, as is inference. It's likely architectural changes that are making their inference cheaper (MoE vs dense models for one), not sure if they've gotten any good methods for making training cheaper.

[–] humanspiral@lemmy.ca 2 points 16 hours ago

There is zero proof of distillation. Minimax 2.7 development was surrounded by moderate use of Claude. M3 is their latest generation, and pretty solid, but its performance cannot be attributed solely (or even 5%) to distillation, and that is only lab that has been accused of significant API use. These claims are all 3-4 months old by now, and Anthropic blocked China access after publishing the accusations. Repeated BS is BS from losers trying to lobby for support.

[–] dil@lemmy.zip 1 points 18 hours ago

Deepseek v4 flash as it is now is perfectly fine for the basic agent stuff I use it for which is basically remote controlling my computer from a distance without using remote desktop. Not sure id be effected significantly if they blocked off their future models.

[–] dil@lemmy.zip 0 points 18 hours ago

Hop on rednote, they aren't fans of their own AI and were collectively throwing fits when claude got banned there.