this post was submitted on 19 Jul 2026
44 points (92.3% liked)

Technology

42961 readers
310 users here now

This is the official technology community of Lemmy.ml for all news related to creation and use of technology, and to facilitate civil, meaningful discussion around it.


Ask in DM before posting product reviews or ads. All such posts otherwise are subject to removal.


Rules:

1: All Lemmy rules apply

2: Do not post low effort posts

3: NEVER post naziped*gore stuff

4: Always post article URLs or their archived version URLs as sources, NOT screenshots. Help the blind users.

5: personal rants of Big Tech CEOs like Elon Musk are unwelcome (does not include posts about their companies affecting wide range of people)

6: no advertisement posts unless verified as legitimate and non-exploitative/non-consumerist

7: crypto related posts, unless essential, are disallowed

founded 7 years ago
MODERATORS
top 18 comments
sorted by: hot top controversial new old
[–] p03locke@lemmy.dbzer0.com 15 points 20 hours ago (1 children)

And the latest Kimi is better than Fable 5. Another plane has indeed hit the Anthropic towers.

[–] communist@lemmy.frozeninferno.xyz -1 points 19 hours ago (1 children)

It does not seem to be better except in terms of cost

[–] p03locke@lemmy.dbzer0.com 15 points 19 hours ago (2 children)

It does not seem to be better

It does.

except in terms of cost

Except cost. Except open source. Except the entire fucking global economic trillion dollar US AI model.

Those are very big exceptions.

(Also, I'm mostly stealing from Yσɠƚԋσʂ's post.)

That's one benchmark that they focused on, but having double checked, you're right, I was thinking of this one where it lags behind gpt 5.6 https://deepswe.datacurve.ai/

[–] dgriffith@aussie.zone 4 points 16 hours ago* (last edited 5 hours ago) (2 children)

I'm just waiting for the closed-source AI industry to have their Black Swan moment.

Something's going to come out of left field from the open AI community and all that investment in proprietary models and mega-compute is going to be rendered useless.

[–] eleitl@lemmy.zip 1 points 19 minutes ago

You still need the mega-compute. Even for inference, 1.5 terabytes of RAM in modern servers isn't cheap.

[–] p03locke@lemmy.dbzer0.com 6 points 15 hours ago

It could be TurboQuant or something like it. The biggest detraction to local LLM models is being able to close the gulf between obscenely-expensive 512GB NPUs, to house the 230GB uncompressed models (+ context), and more common 24GB GPUs. Quantized 15-18GB models are already working pretty well, but context size is still a bit of a problem.

Of course, the whole industry need to ramp up memory production and wrestle duopolies from the few that can make the raw silicon. It was pretty fucking pathetic that parts of the PC industry decided to leave these silicon processing weaknesses in various places. Large corps could have easily jumped into the industry and made bank in the long-term, but that would require not funneling into short-term quarterly profit bullshit.