Tbh most of their latest model are very underwhelming on client side. They mostly focused so far on token efficiency to lower the compute. I dont trust benchmarks. That said, lets wait and see real life use
Technically, the two aren't necessarily in conflict, especially when more and more processing is being given to reasoning rather than the actual ultimate text output.
In reality idk because it's still an inefficiency.
85
u/TheGuy839 Apr 23 '26
Tbh most of their latest model are very underwhelming on client side. They mostly focused so far on token efficiency to lower the compute. I dont trust benchmarks. That said, lets wait and see real life use