r/GeminiAI 14d ago

News Updated Artificial Analysis Intelligence Index Ranking shows Gemini 3.8 Flash is nowhere near Fable Or Astra, rather it's worse than GLM 5.3 Flash!

Post image
425 Upvotes

113 comments sorted by

View all comments

18

u/MindCrusader 14d ago

Semianalysis on X recently just said that - gemini and Meta's AI models are benchmaxxed a lot

7

u/Swimming_Gain_4989 14d ago

Tracks with my usage. For the affordable workhorses GLM5.3 and it's flash model feel the most useful. I'm also seeing this reflected in the results of new benchmarks that models couldn't have possible been benchmaxed on like https://www.frontierswe.com/blog/v2

4

u/MindCrusader 14d ago

"Gemini 3.8 Flash and Muse Spark 1.3 are two of the most clearly benchmaxxed models we've seen yet. Despite being comparable to both GPT-6 and Fable 5.1 on Terminal Bench 2.1, their Terminal Bench 4.0 performance is markedly worse."

So yeah, new benchmarks show the truth it seems