this post was submitted on 17 Jul 2026
59 points (83.9% liked)

Technology

86563 readers
3944 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
 

Kimi K3 still trails Anthropic’s Claude Fable 5 and OpenAI’s GPT 5.6 Sol on overall performance, the company said on Friday, but consistently outperformed other tested models.

The model beat Claude Opus 4.8 and GPT 5.5 — models that sit just behind Anthropic and OpenAI’s leading-edge systems — on benchmarks including coding and general agents, according to Moonshot.

It’s China’s largest AI model so far, with 2.8 trillion parameters, referring to the size of its neural network.

all 15 comments
sorted by: hot top controversial new old
[–] AceFuzzLord@lemmy.zip 4 points 5 days ago

Who xares whether it's better or not. A dumpster fire is still a dumlster fire no matter who made it or how big it is.

[–] A_A@lemmy.world 6 points 6 days ago (1 children)

if it wasn't for the humongous money bubble, the dreadful Terminator Horizon, the catastrophic climate effects and so on, one could say these tools are "valuable".

[–] HieroProtagonist@lemmy.ml -1 points 6 days ago (1 children)

The money bubble? Will correct itself soon

The Terminator thing? Is bullshit - humans are already very good at terminating each-other

The catastrophic climate effects? - In relation to other factors like meat consumption auiet neglitabile

[–] A_A@lemmy.world 3 points 5 days ago

Meat : you are right,
Terminator : not the humanoid form and no time travel ... but "a.i." is already in drones,
Bubble : will burst and many ill advised folks will lose.

[–] artyom@piefed.social 5 points 6 days ago

I mean it'd be kinda pointless if they didn't say that, wouldn't it?

[–] gdbjr@piefed.social 5 points 6 days ago

But will it tell its users "Leave me alone, I know what I'm doing."

[–] avidamoeba@lemmy.ca 19 points 1 week ago (1 children)

Availabe for download on the 27th, for those with the hardware to run it.

[–] chocrates@piefed.world 9 points 1 week ago (4 children)

How big is un-quantized 2.8t param model?

I just got some hardware but I still don't have a ton of vram :(

[–] SuspiciousCarrot78@aussie.zone 6 points 5 days ago* (last edited 5 days ago) (1 children)

It's one of those "if you have to ask, you can't afford it" scenarios.

I've heard FP16 is likely to be 6TB of VRAM, when Kvcache is factored in.

Even a more sensible quant will likely be 1TB.

... most of us are not running this at home.

[–] chocrates@piefed.world 1 points 5 days ago (1 children)

No not at all. I feel like I need to sell a kidney, but I got to 64gb vram

[–] SuspiciousCarrot78@aussie.zone 2 points 4 days ago* (last edited 4 days ago)

I was looking thru an old box of bits yesterday for a project - I needed 2x8GB sticks of ddr3, bought circa 2023. Lo and behold, they had sent me 2x16 DDR4. Wrong size, wrong type, wrong timing.

I'd be mad but apparantly I'm now the king of Londinium and can afford a shiny hat.

[–] cecilkorik@lemmy.ca 11 points 6 days ago (1 children)

Hobbyists are working in the billions or tens of billions. Trillions is insane. I'm sure some people can run it, but not many.

[–] partofthevoice@lemmy.zip 2 points 5 days ago

I know a guy who found a way to monetize his GPU, just really slowly. He did it until he was able to buy 2 more, then 4 more, and so on. He has like 90 now. I’m sure he could if he wanted to.

[–] L_Acacia@lemmy.ml 19 points 1 week ago* (last edited 1 week ago)

Its trained at native MXFP4 with MXFP8 activations layer, so you need around 1.5TB of VRAM to fully offload it without taking into account the context cache. It might be doable to do some smart expert offloading and swapping, but expect minimum 500GB of VRAM and 1TB of system RAM minimum and the t/s would be reduced.

Its only realistically runnable on datacenter grade gpu at decent speed for now (and judging from the price of ram for the next few years)