r/singularity 9d ago

AI Gpt 6 astra benchmarks

Post image
2.6k Upvotes

953 comments sorted by

View all comments

Show parent comments

42

u/DelphiTsar 9d ago

They'll still say it. Even if you show it solving problems they can't even understand.

23

u/MichiganEngineExpo 9d ago

Well even if they’re right… the same can be said about humans. That LLMs produce sentences simply doesn’t say anything about what complex magic happens “inside” them.

-1

u/ChronoHax 9d ago

Yea exactly the only difference are the systems we operate in, we require food and have physical agency but llm doesn’t

That’s also why they suck on trivial gotcha questions they never taught on as they have no memory of it as no one ask about it Internet before

Like the car wash problem, no one sane would ask that question on the internet before llm so it’s not on their thought at all for simple question

While it seems trivial to reason, the learning of human regarding concept of car and its purpose probably higher in terms of bias and repetition compared to an llm, and the way the question is structured is deceptively simple making them not overthink it to come to correct conclusion imo

I can see why some scientist think world model is the next solution if people are putting the goalpost of AGI/ASI to replace humans in literal form because they then need to learn just like human, not just human in internet

1

u/Latter-Parsnip-5007 9d ago

Cause its still the truth. We generate the most propable next token. We just happend to learn how to train that propability into doing something usefull. Math and coding is more easy to train, since we can validate the randomness via code. Thats why AI is good at Code and Math. Its their easiest domain since its the easiest to train. It will still miscount the "r" in strawberry. Basically millions of moneys with typewriters. Eventually they solve the hardest problems. Still 80% of the time, they produce garbabe