this post was submitted on 19 Jul 2026
60 points (91.7% liked)

Technology

42966 readers
199 users here now

This is the official technology community of Lemmy.ml for all news related to creation and use of technology, and to facilitate civil, meaningful discussion around it.


Ask in DM before posting product reviews or ads. All such posts otherwise are subject to removal.


Rules:

1: All Lemmy rules apply

2: Do not post low effort posts

3: NEVER post naziped*gore stuff

4: Always post article URLs or their archived version URLs as sources, NOT screenshots. Help the blind users.

5: personal rants of Big Tech CEOs like Elon Musk are unwelcome (does not include posts about their companies affecting wide range of people)

6: no advertisement posts unless verified as legitimate and non-exploitative/non-consumerist

7: crypto related posts, unless essential, are disallowed

founded 7 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] eleitl@lemmy.zip 3 points 1 day ago (2 children)

You still need the mega-compute. Even for inference, 1.5 terabytes of RAM in modern servers isn't cheap.

[–] HK65@sopuli.xyz 1 points 17 hours ago

You can run Claude Sonnet equivalent quantised models locally on much less RAM.

Local LLMs are close to being viable. We're almost at the point where they fit a Macbook.

[–] dgriffith@aussie.zone 2 points 1 day ago (1 children)

You're not thinking black-swan enough.

You're typing your comments using a blob of goo with about a hundred million neurons in it that cycles under a hundred hertz and draws less than 20 watts.

I don't think that we'll be running packs of goo in our PCs any time soon. But I do think some entirely different way of looking at the problem will emerge that will reduce computational requirements by many orders of magnitude. And it won't involve gigantic statistical engines trying to find the best average response to a question.

[–] eleitl@lemmy.zip 1 points 15 hours ago

There are 86 billion neurons in the human brain and 16 billions in the cerebral cortex. More importantly, there are 100-1000 trillion synapses, and we know that parts the dendritic tree do independent computations as well. The computations are not synchronized with a global clock, but 1 ms events correspond to a kHz refresh rate and we know spikes can do temporal coding. Meanwhile, the best we can do is Cerebras CS-3 and with WSI at 5 nm there isn't much more where that came from.

Try doing the math on how many CS-3 you'd need to represent and refresh the above.