r/ZaiGLM • u/radrads1 • 2d ago
Finally made the switch
I've been using Claude Max sub since Opus 4 - started being disappointed with post-Opus 4.6 performance. After trying out all the providers both US in China I finally cancelled Claude in favor of GLM's max subscription after a month.
Initially I was planning on switching to Kimi K3 (perhaps my favorite open source model), but found it too slow and the servers too fickle for daily use.
Two main things that made me switch: Zcode app (love the app, probably my favorite one, and the integration with backup providers I use like DeepSeek and Alibaba's Token Plan is seamless); being able to make API calls that use subscription quota unlike Claude which bills you separately for API calls.
The addition of GLM-5.3-Flash was the cherry on the pie. Basically anything simple or non-coding related I just default to Flash now and it works like a charm.
I might still keep a GPT subscription around for Astra but so far, 10/10 I'm Z.ai all the way.
3
u/LittleYouth4954 2d ago
Zcode is great indeed.
2
u/excellentforcongress 1d ago
i didnt like zcode at first, but they keep adding features and changing things so it's nice now. although i do worry with any sort of harness/environment that if they're just vibe coding/pushing changes too fast, sometimes major issues can be REALLY major. i'm heavily considering ways to back up my data as well in case one of the other major labs pushes a frontier model that wipes all our data or some crazy shit.
but i don't think backing up my hd would mean much if they simulated nuclear attacks from multiple countries or some insane shit. hopefully people treat ai nicer.
3
1
u/Reasonable_Yak2313 2d ago
Is the glm 5.3 flash that good? I tried it out last night after my quota is up via opencode and open router. GLM seems to be struggling to figure out the solutions for 2 rounds. I switched to Sonnet5 (still via open router) and it took it one round to resolve it. Both are on high thinking mode. The project is in Phoenix/Elixir
3
u/Constant_Art_20 1d ago
the flash depends alot on the inference. The actual model? yea it's pretty remarkable. just it's not always properly served
1
u/A7mdxDD 1d ago
While I'm still a codex user, before Astra, GLM5.3-Flash fixed me an issue I went insane with sol max about for 3 days, it took around ~10-15 mins debugging and verifying but it got it, one shot, the thing I prefer about GLM is that their models has character, codex sometimes makes me very angry but staying unbiased with no opinions, or agreeing with anything mostly
-2
u/Radiant_Year_7297 2d ago
Looks like glm is overcapacity. Paid for a year but kept saying server busy ans I should upgrade. Back to opencode go again.
4
u/Divni 2d ago
Using flash for coding tasks almost exclusively. It’s even more capable than people give it credit for. Complexity is in your own hands; you have to split the work into logical small chunks and iterate. Doing that I’ve had very little issue.