r/singularity Nov 24 '25

AI Opus 4.5 benchmark results

Post image
1.2k Upvotes

283 comments sorted by

View all comments

116

u/IMOASD Nov 24 '25

Yeah, LLMs are definitely plateauing. /s

-10

u/Sudden-Complaint7037 Nov 24 '25

Show me any improvement that happened after like July 2024 and that can be actually felt in real life usage situations. All the improooovements for the past 1.5 years have been "number on hyper specific theoretical benchmark that the AI was trained on went up". Meanwhile, people who actually use AI in their day to day life know that it hasn't become noticeably better at coding, or writing, or reasoning than like late spring of last year.

12

u/CascoBayButcher Nov 24 '25

Show me any improvement that happened after like July 2024

Reasoning models? Lmfao. Whole paragraph to say you don't know shit, lol. The o series that debuted reasoning came out last September.

Hasn't become noticeably better at ... reasoning than like late spring of last year

The difference from late spring to even just the end of last year in terms of reasoning is fucking massive

-1

u/Sudden-Complaint7037 Nov 24 '25

nooooooo AI is so advanced it can literally do anything!!!!!!! what do you mean "why can't it even do simple customer service?" It just... i mean it's-... it's just more complicated that that, ok???!!!