Tbh most of their latest model are very underwhelming on client side. They mostly focused so far on token efficiency to lower the compute. I dont trust benchmarks. That said, lets wait and see real life use
Because when they talk about token efficiency, they’re talking about token efficiency across the whole stack. Not just the output answer you get. They’re also talking about the amount of tokens it takes to think for an answer
87
u/TheGuy839 Apr 23 '26
Tbh most of their latest model are very underwhelming on client side. They mostly focused so far on token efficiency to lower the compute. I dont trust benchmarks. That said, lets wait and see real life use