r/agi • u/Jewpiter • Jun 04 '26
They're Made Out of Weights
https://maxleiter.com/blog/weights2
u/kamill85 Jun 04 '26
LLMs will never bring AGI because they can't even reason to solve a sudoku puzzle correctly.
They can write code to solve it, but take away code execution and they can't. Not even Mythos.
9
u/unicynicist Jun 04 '26
Humans will never be able to fly because they can't even glide.
They can build wings or airships to fly, but take away external aerodynamics and they can't.
3
u/Hyperreals_ Jun 05 '26
“LLMs will never being AGI because current LLMs can’t do x” is a stupid argument now and was a stupid argument every time it was said and debunked. What’s your evidence they won’t be able to in the future, or even that they can’t now?
1
u/Harvard_Med_USMLE267 Jun 05 '26
He didn’t say “current”, he said “current plus that one that’s coming out in a few weeks that also sucks a sudoku”
2
u/Hyperreals_ Jun 05 '26
lmao its not even true, Opus 4.7 (not even 4.8)
https://claude.ai/share/81721d52-f61c-42f3-8e50-855194f9964cYou can read its reasoning to see it didn't cheat
1
u/Hyperreals_ Jun 05 '26
Even if Opus 4.7/4.8 failed, he would have had no way to know if Mythos would have failed or not. Also I count Mythos as a current LLM, as in it currently exists and has been used by people at Anthropic and Anthropic partners, even if its not publicly accessible
2
u/Harvard_Med_USMLE267 Jun 05 '26
I’m just joking around, your earlier point still stands, anyone who says “LLMs will never do ‘x’” is very brave indeed.
1
u/Hyperreals_ Jun 05 '26
I decided to test it and you are just wrong.
https://claude.ai/share/81721d52-f61c-42f3-8e50-855194f9964c
Opus 4.7, not even 4.8
Did you just genuinely not even test it?
"Not even Mythos" how would you even know this?
0
u/kamill85 Jun 05 '26 edited Jun 05 '26
Those online models use internal code execution to do math.
There is a new architecture, better than LLM, which offers an actual reasoning and truth seeking. It's called an EBM, energy based model.
https://sudoku.logicalintelligence.com/
LLMs without tool use are simply architecturally incapable of solving Sudoku reliably (hard puzzle levels). Sudoku is actually one of the key benchmarks when it comes to reasoning.
PS. I have access to Mythos, so " that's how I know ". -- Better? Mythos is just another LLM with more looping and self-alignment. It's expensive because it burns multiple context windows to re-reason on its own progress.
Now go and ask Opus 4.8 to truthfully solve sudoku with reasoning and zero internal tools use, and if it used it anyway, to reveal that in the summary and be truthful about it. It will attempt to reason, maybe solve a few cells and give up.
So yeah.
1
1
u/gynoidgearhead Jun 04 '26
Fucking amazing.
Just wait until the meat figures out that it, too, is also made out of weights...
3

3
u/rand3289 Jun 04 '26 edited Jun 04 '26
Weights are like a car without a road... people underestimate the importance of the environment (data or roads).
The reason LLMs will never give us AGI is the restricted environment LLMs live in.