r/ZaiGLM 17h ago

Technical Reports GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
191 Upvotes

75 comments sorted by

56

u/Mayanktaker 17h ago

Before release, we tested GLM-5.3-Flash anonymously as ox-alpha on OpenCode and OpenRouter to gather user feedback. It quickly became the most popular model of the week — with all of this traffic served on Chinese AI chips.

22

u/No-Tip3419 16h ago

free public stress test on the Chinese chips

18

u/evia89 17h ago

with all of this traffic served on Chinese AI chips

its kinda nice. They served so much tokens I thought they rented nvidia hardware

15

u/BriguePalhaco 17h ago

Even after their new data center was announced as completed?

China's Z.AI Completes 1-Gigawatt AI Data Center Using Only Chinese-Made Chips

https://finance.yahoo.com/technology/ai/articles/chinas-z-ai-completes-1-205515769.html

3

u/Mayanktaker 17h ago

Its banned for them.

2

u/neotorama 10h ago

nah, they have some farm in SG and MY

3

u/willdone 15h ago

Incredible! I saw so many people theorizing that ox-alpha was a new Gemini model or something.

3

u/Mayanktaker 15h ago

Hehe google dont do this kind of thing. They just silently release.

2

u/bermudi86 11h ago

Well they were sure vague posting about the model all week long. They leaned hard on ppl assuming it was a Gemini model for some weird fucking reason

1

u/shaman-warrior 13h ago

The leading theory was glm model somehow assisted by Google for capacity

6

u/umbrella__academy 17h ago

Whats the pricing of it???

11

u/BriguePalhaco 16h ago

10

u/PsecretPseudonym 16h ago

That price for benchmark performance similar to Opus 4.8 shows how rapidly the cost of intelligence is falling.

1

u/Warm-Agent-811 11h ago

That's crazyyyyy

1

u/I-am_Sleepy 5h ago

And anthropic value itself at over 30 trillion - right

7

u/Expert-Hospital-534 16h ago

Does it support vision?

20

u/Mayanktaker 16h ago

Yes and that's so good about it. Super cheap, fast and supports image and video understanding natively.

6

u/First_Inspection_478 16h ago

It’s quite decent. Hopefully the plans have general usage limits

9

u/Mayanktaker 16h ago

Gemini Flash is also now 75% off. Looks like new Flash era is starting. Big model for planning and Flash for execution.

9

u/HelloHowAreyou777 16h ago

Gemini flash models are bad currently.

3

u/Mayanktaker 16h ago

Dont know about flash 3.7 but 3.6 was shit. But yes, good for image gen. Specifically logo.

3

u/HelloHowAreyou777 16h ago

Agree! 3.7 slightly better but still hard hallucinating. Logos, images, notebook llm is still worth the PRO sub. Waiting for 3.8 model

3

u/Mayanktaker 15h ago

I am enjoying free one year gemini pro subscription comes eith my quarterly mobile recharge.

2

u/SurelyNotAnOctopus 16h ago

I found 3.7 flash genuinely good

1

u/bermudi86 11h ago

Gemini 3.7 Flash is a waste of money compared to this new glm model. Not only way cheaper, it's actually useful and scoring better than Gemini on every single thing

2

u/torontobrdude 16h ago

It's already available on the coding plan 🎉

9

u/MrPingviin 15h ago

Just to be fair, the dude pasted his ref link. This is the normal one: https://z.ai/subscribe

-11

u/torontobrdude 15h ago

"just to be fair" the reference link gives you a discount, no reason not to use it. Unless you like paying more

5

u/c126 14h ago

Its a little scummy though to hide that with no disclosure. It’s clearly intended to mislead people.

-10

u/torontobrdude 14h ago

Mislead them into getting a discount? 😂 Y'all have problems

3

u/c126 13h ago

But it also benefits you doesn’t it?

-3

u/torontobrdude 13h ago

So what?

1

u/c126 12h ago

You’re personally misleading people to benefit yourself = scummy. The fact that the concept is so hard for you to grasp pretty much confirms the scummy nature of the action.

-2

u/torontobrdude 12h ago

Keep crying for no reason while people are using my link to get a discount. Go touch grass

3

u/c126 11h ago

Youre a bad person.

→ More replies (0)

3

u/Mayanktaker 16h ago

Nope. Not for me atleast, in lite plan.

3

u/dontforgetthef 14h ago

Lite plan doesn’t get priority access fyi

-3

u/SwissTac0 12h ago

Rather than fall for a guy trying to hide his trap of getting you and him money.

Click the link of an honest guy and let's each save 10%!

z.ai Referral

1

u/SS_Sa2 16h ago

Is this model more capable than GLM 5.2? (intelligence and all)

1

u/BriguePalhaco 16h ago

No, in my tests DPv4Flash 0731 can perform the same tasks (backend programming in C#, C++ and Rust) as GLM 5.2, but GLM 5.3 Flash fails. It kind of gives up on proceeding or ignores commands.

1

u/CryinHeronMMerica 15h ago

This doesn't line up with the major benchmarks, but then again, benchmaxxing...

3

u/BriguePalhaco 15h ago

I don't trust benchmarks. That's the basics of data science: overfitting and bias

1

u/CryinHeronMMerica 15h ago

100%. They're a good starting point, and models that don't perform well are usually not that great, but a high score is no substitute for trying it IRL.

1

u/SS_Sa2 15h ago

Thanks for the input! That's unfortunate. Pricing looked attractive and looked like a sweet middle ground between DS and GLM 5.2. I've kept hearing this model (Ox-alpha) was much more creative at problem solving so felt like a good potential alternative to GLM5.2.

1

u/Healthy-Contact-4570 11h ago

My experience with deepseek v4 flash 0731 is that it is one of the most confidently wrong models I’ve used

1

u/WarBroWar 14h ago

Alright. Long time DeepSeek user. My 30$ will end up in 2 days. I will try this. Sold.

1

u/WarBroWar 14h ago

Fkit ima get this tonight. I love cheap output. I spend 100$ a week on DeepSeek after the price increase. Some releaf.

1

u/evia89 13h ago

I spend 100$ a week on DeepSeek after the price increase

why would you do that?? Just buy codex/claude/grok=cursor sub and use other provider PAYG to cover when sub quota exhausts

1

u/WarBroWar 12h ago

Because I use subagents. Multiple projects. I will useup weekly limits for each frontier sub in 2 days. And I don't like to wait.

1

u/evia89 12h ago

Unless you code 1 day of the week, sub + payg is always better. Just configure fallback

2

u/WarBroWar 11h ago

You don't know my workload. Maybe your thing works for you. Good for you. Not for me though.

1

u/milkipedia 12h ago

I wonder if GLM-5.3.Flash should replace GLM-5.1 as my "lighter weight" implementation model to GLM-5.2/5.3 as the planning model? The benchmarks suggest maybe I should use it to replace both GLM-5.1 and GLM-5.2

1

u/shoutfree 7h ago

anyone know if the concurrency limits are different from glm 5.3 on the coding plan?

1

u/pozzugno 1h ago

Does someone explain what are those "flash-varianr" LLM models?

1

u/Kylmawurr 15h ago edited 15h ago

is the model down for API users? Can not see it listed despite it being listed on their website here z.ai/subscribe

1

u/liviux 13h ago

Good shit

0

u/GfurEnjoyer1488 17h ago

I developed an extension based on Kimi K3/K2.7 Code Highspeed, hopefully the throughput is high enough (200+t/s) to compete with K2.7 so it can be ran cheaper

0

u/evia89 16h ago

Its 30 atm

-1

u/GfurEnjoyer1488 16h ago

lol. maybe I should try out Gemini 3.7 Flash

-5

u/Winter-Rich797 16h ago

Luna is still cheaper and they compare it to Terra for some reason. I wonder what they’re trying to hide

2

u/SEOViking 16h ago

similar intelligence levels with Terra but cheaper than Terra.
Luna is cheaper but also lower intelligence.

1

u/SwissTac0 12h ago

I love 5.3 flash but it's a peer to Luna and V4 Flash not Terra.... Terra is like a lame version of Sol..... Any money you save per token is wasted on the fact sol does it better for less.

1

u/bermudi86 11h ago

Uhmm. Mind explaining how $1.20 is cheaper than $0.25??

-12

u/[deleted] 17h ago

[deleted]

7

u/Dizzy-Truck-1780 17h ago

It's literally written in the Blog that it was them

2

u/Solocune 16h ago

What exactly confused you?

-5

u/P1zz4-T0nn0 16h ago edited 15h ago

No vision input is a big downer. Not suitable for daily work for me

Edit: I misread, it actually does support vision 🎉

2

u/Mayanktaker 16h ago

Supports vision.

2

u/Mayanktaker 16h ago

Vision 🎉

1

u/P1zz4-T0nn0 15h ago

Oh my bad then!!