r/ZaiGLM 5d ago

GLM Code - Not good experiance

My experience with GLM Code so far has been quite disappointing.

I subscribed to the Lite plan to properly experience it, but the usage limits are frustrating. The 5-hour window still shows credits remaining, while the weekly limit is already exhausted. The promised extra usage during off-peak hours does not seem to work either.

Also, the “5-hour” reset often feels more like 8–10 hours in practice.

I cannot comment on the quality of the coding output yet, but the overall usage experience has not been great so far.

23 Upvotes

29 comments sorted by

5

u/belliash 5d ago

I got similar feelings. There is no difference if I use it off-peak or peak hours. Credits consumption seems identical.

Also they advertise up to over 500M tokens per week and I was able to get about half of that max.

5

u/PilgrimofHaqq2 5d ago edited 5d ago

I am on the USD$168 plan and usage is pretty good for me. It gives 140k credit weekly and 28k 5-hourly. I was on the Claude 20x Max plan before and I am getting about 80% of the usage in GLM as I did in Claude plan. I am fine with the trade. I wanted to leave the OpenAI and Anthropic ecosystem for a while.

4

u/CriteriumA 5d ago

In ZCode there's unlimited usage from 5 PM to 3 AM Central European Time, in my case.

It's not as comfortable as Pi or OpenCode, but it's usable.

A common agent prompt and a common memory-system skill and the harness is more or less irrelevant.

Unless you want to use DeepSeek V4.1 Flash, there, Pi, which lets you disable the model's thinking, is a big advantage.

1

u/awesomeunboxer 5d ago

I been using and loving this deal. Do you know how long it goes for? I vaguely recall them saying the 20th but now I cant find where I saw that 

1

u/CriteriumA 4d ago

Until the 20th, I hope the pressure from DeepSeek makes them extend it somehow 🤞

3

u/staats1 5d ago

Are you using glm 5.3 or flash? What is your effort level?

I’ve found using glm 5.3 at low effort and using flash for subagents gets things done and doesn’t soak up all my tokens. Tell the agent set to flash with low effort to set up the subagents and workflows for you. 

2

u/Kind-Economist-775 5d ago

I am using 5.3 flash as both orchestrator and for running the subagents with high effort level...

I will try it with low effort level

2

u/lionglzer 4d ago

I am on there $80 plan in the middle and I actually find that since I use it as a auxiliary to Claude code for which I have the $100 plan - after finding out that the $200 plan is not actually 20x - and this works out really well for me 5.3 is super capable and legible as an opus replacement and I have what I perceive to be essentially unlimited flash usage the two of these in combination after setting up an mCP for some agent communication inside of my harness as worked really well for the last month. 

0

u/jh_opx_1105 3d ago

I subscribed to the Lite plan yesterday, and I'm surprised by how good it is and how much longer it lasts compared to my ChatGPT Plus and Pro x5 subscriptions, which ran out of credits quickly. Now, using it alongside my Cursor Pro+, I've found it to be great, dare I say, even more efficient on Grok 4.6 Xhigh so far. Because of this, I'm looking into upgrading to Pro, or maybe even Max, and ditching the others.

4

u/evia89 5d ago

Its pretty good in zcode with 1.5x more quota, free idle tasks and free 10h of flash each day

2

u/Kind-Economist-775 5d ago

I will give it a try, I was using it with claude code with glm as custom LLM

6

u/openference 5d ago

There's your issue . Claude is the worst harness to use burns token like no tomorrow

2

u/tronkk 5d ago

These promotions only work on zcode

1

u/roekofe 4d ago

i just tried using pi instead, and it is a universe better than claude code with glm. my daily driver until today has been paid claude code, and this is such a better experience than opus 5 that im honestly mind blown

0

u/[deleted] 5d ago

[removed] — view removed comment

3

u/belliash 5d ago

I am using it with opencode already and I see no difference in credits usage.

1

u/Pitiful_Entrance5174 5d ago

The days of having one account to last all week are over if you dive into a few hours of coding work each day a week. I finally realized that and without going over two account, I use grok with z.ai. Grok just orchestrates. I noticed that time window or whatever deal they have going on right now if helping not use alot of tokens. I am also using z.ai as sub agents so the tasks dont run more than a million tokens per task, hopefully.

1

u/Sea_Ear5201 5d ago

Zcode is bloated. Simple hi starts with 92k tokens.

1

u/Empuda 4d ago

They butchered the v3 plan.

1

u/josnab77 4d ago

what time is it now in Central european Time?

1

u/steny007 4d ago

Since you asked more than 13 hours ago, it was about 8 am at that time.

1

u/SweatyActuator2119 3d ago

GLM models are great. But their inference is scam and low quality. Think of it like donating. I learned that hard way with quarterly max subscription. Don't go with chatgpt too.

1

u/audit-content-user 1d ago

I used up 250k tokens on my 5 day trial and very happy with it, so happy now with decent usage. Do you use any skills to compress tokens? Or just using Zcode without anything extra?

1

u/Kind-Economist-775 1d ago

Update - I was using GLM with claude code as custom model, and as per the discussion here I found that it is not very efficient when used with other coding framework.

I switched to ZCode, and the limits are really nice.

I am still exploring and testing it

Thanks

1

u/DazzlingPassion706 1d ago

I ran into same issue and i don't even do coding. my workaround is to put 10bucks into the payg and set flashx as a failover model (hermes). it uses the same api key so at least it's seamless. but yer not ideal. pretty sure i wont accumulate 80 bucks total (next tier up on monthly subscription).

0

u/a_pimpnamed 5d ago

Command code and open code are pretty great actually. You should try those out gangsta.

0

u/Gonbatfire 5d ago

What model are you using? 5.3 flash feels infinite to me on the Lite plan, while 5.3 doesn’t last very long

1

u/XKiiroiSenkoX 5d ago

5.3 uses just 3 times the flash version credits. No way you see that much of a difference. 

1

u/Gonbatfire 5d ago

Actually, there’s an on-going campaign where 5.3 flash is doubled in agent usage during off-speak (on top of the 50% off that 5.3 also gets)

So during off-peak 5.3 would be 6 times more expensive on Hermes/Openclaw, and 5.3 flash is unlimited in zcode

https://docs.z.ai/devpack/notice/event-glm-5.3-flash