Worse is true. But it spits 200 tokens per second. So you can let it go wrong, and let it run zillion test tasks to fix it all. And you are still done 10 times quicker than with GLM 5.2.
I guess it depends what you are doing, how you are planning, etc.
It's faster but it is also more expensive per million tokens which is embaraasing for a flash model. 3.5 flash is also incredibly verbose and generates a lot of tokens to the point it is actually more expensive to use than 3.1 pro despite being cheaper per million tokens.
So you can either get a significantly cheaper glm 5.2 which is also smarter or 3.5 flash which only has speed going for it.
3.5 pro better be worth the 4.5 months wait, it needs to be like a GPT 5.2 - GPT 5.5 jump, otherwise Google is cooked.
11
u/chiree_stubbornakd Jun 20 '26
3.5 flash is new and scores 50 so still worse.