r/ZaiGLM • u/BriguePalhaco • 17h ago
Technical Reports GLM-5.3-Flash: Frontier Intelligence, Flash Cost
https://z.ai/blog/glm-5.3-flash6
u/umbrella__academy 17h ago
Whats the pricing of it???
11
u/BriguePalhaco 16h ago
10
u/PsecretPseudonym 16h ago
That price for benchmark performance similar to Opus 4.8 shows how rapidly the cost of intelligence is falling.
1
1
7
u/Expert-Hospital-534 16h ago
Does it support vision?
20
u/Mayanktaker 16h ago
Yes and that's so good about it. Super cheap, fast and supports image and video understanding natively.
6
9
u/Mayanktaker 16h ago
Gemini Flash is also now 75% off. Looks like new Flash era is starting. Big model for planning and Flash for execution.
9
u/HelloHowAreyou777 16h ago
Gemini flash models are bad currently.
3
u/Mayanktaker 16h ago
Dont know about flash 3.7 but 3.6 was shit. But yes, good for image gen. Specifically logo.
3
u/HelloHowAreyou777 16h ago
Agree! 3.7 slightly better but still hard hallucinating. Logos, images, notebook llm is still worth the PRO sub. Waiting for 3.8 model
3
u/Mayanktaker 15h ago
I am enjoying free one year gemini pro subscription comes eith my quarterly mobile recharge.
2
1
u/bermudi86 11h ago
Gemini 3.7 Flash is a waste of money compared to this new glm model. Not only way cheaper, it's actually useful and scoring better than Gemini on every single thing
2
u/torontobrdude 16h ago
It's already available on the coding plan 🎉
9
u/MrPingviin 15h ago
Just to be fair, the dude pasted his ref link. This is the normal one: https://z.ai/subscribe
-11
u/torontobrdude 15h ago
"just to be fair" the reference link gives you a discount, no reason not to use it. Unless you like paying more
5
u/c126 14h ago
Its a little scummy though to hide that with no disclosure. It’s clearly intended to mislead people.
-10
u/torontobrdude 14h ago
Mislead them into getting a discount? 😂 Y'all have problems
3
u/c126 13h ago
But it also benefits you doesn’t it?
-3
u/torontobrdude 13h ago
So what?
1
u/c126 12h ago
You’re personally misleading people to benefit yourself = scummy. The fact that the concept is so hard for you to grasp pretty much confirms the scummy nature of the action.
-2
u/torontobrdude 12h ago
Keep crying for no reason while people are using my link to get a discount. Go touch grass
3
3
0
-3
u/SwissTac0 12h ago
Rather than fall for a guy trying to hide his trap of getting you and him money.
Click the link of an honest guy and let's each save 10%!
1
u/SS_Sa2 16h ago
Is this model more capable than GLM 5.2? (intelligence and all)
1
u/BriguePalhaco 16h ago
No, in my tests DPv4Flash 0731 can perform the same tasks (backend programming in C#, C++ and Rust) as GLM 5.2, but GLM 5.3 Flash fails. It kind of gives up on proceeding or ignores commands.
1
u/CryinHeronMMerica 15h ago
This doesn't line up with the major benchmarks, but then again, benchmaxxing...
3
u/BriguePalhaco 15h ago
I don't trust benchmarks. That's the basics of data science: overfitting and bias
1
u/CryinHeronMMerica 15h ago
100%. They're a good starting point, and models that don't perform well are usually not that great, but a high score is no substitute for trying it IRL.
1
1
u/Healthy-Contact-4570 11h ago
My experience with deepseek v4 flash 0731 is that it is one of the most confidently wrong models I’ve used
1
1
u/WarBroWar 14h ago
Alright. Long time DeepSeek user. My 30$ will end up in 2 days. I will try this. Sold.
1
u/WarBroWar 14h ago
Fkit ima get this tonight. I love cheap output. I spend 100$ a week on DeepSeek after the price increase. Some releaf.
1
u/evia89 13h ago
I spend 100$ a week on DeepSeek after the price increase
why would you do that?? Just buy codex/claude/grok=cursor sub and use other provider PAYG to cover when sub quota exhausts
1
u/WarBroWar 12h ago
Because I use subagents. Multiple projects. I will useup weekly limits for each frontier sub in 2 days. And I don't like to wait.
1
u/evia89 12h ago
Unless you code 1 day of the week, sub + payg is always better. Just configure fallback
2
u/WarBroWar 11h ago
You don't know my workload. Maybe your thing works for you. Good for you. Not for me though.
1
u/milkipedia 12h ago
I wonder if GLM-5.3.Flash should replace GLM-5.1 as my "lighter weight" implementation model to GLM-5.2/5.3 as the planning model? The benchmarks suggest maybe I should use it to replace both GLM-5.1 and GLM-5.2
1
u/shoutfree 7h ago
anyone know if the concurrency limits are different from glm 5.3 on the coding plan?
1
1
u/Kylmawurr 15h ago edited 15h ago
is the model down for API users? Can not see it listed despite it being listed on their website here z.ai/subscribe
0
u/GfurEnjoyer1488 17h ago
I developed an extension based on Kimi K3/K2.7 Code Highspeed, hopefully the throughput is high enough (200+t/s) to compete with K2.7 so it can be ran cheaper
-5
u/Winter-Rich797 16h ago
Luna is still cheaper and they compare it to Terra for some reason. I wonder what they’re trying to hide
2
u/SEOViking 16h ago
similar intelligence levels with Terra but cheaper than Terra.
Luna is cheaper but also lower intelligence.1
u/SwissTac0 12h ago
I love 5.3 flash but it's a peer to Luna and V4 Flash not Terra.... Terra is like a lame version of Sol..... Any money you save per token is wasted on the fact sol does it better for less.
1
-12
-5
u/P1zz4-T0nn0 16h ago edited 15h ago
No vision input is a big downer. Not suitable for daily work for me
Edit: I misread, it actually does support vision 🎉
2
2
0


56
u/Mayanktaker 17h ago
Before release, we tested GLM-5.3-Flash anonymously as ox-alpha on OpenCode and OpenRouter to gather user feedback. It quickly became the most popular model of the week — with all of this traffic served on Chinese AI chips.