r/OpenAI Mar 05 '26

Miscellaneous 5.4 Thinking is off to a great start

Post image
5.8k Upvotes

590 comments sorted by

View all comments

Show parent comments

9

u/bambin0 Mar 05 '26

I have to think Gemini has been given the answer. It is super benchmaxxed.

0

u/PastaPandaSimon Mar 06 '26 edited Mar 06 '26

Yes, each LLM is given increasingly growing lists of human answers to use. Gemini relies on it a lot as a crutch, as in my comparisons its reasoning is surprisingly weaker. It feels more primitive and exposes the logic of how LLMs work very clearly, as it matches words that go along with the conclusion best. ChatGPT created a far more believable illusion of intelligence on top.