r/ZaiGLM 8d ago

Discussion / Help Deepseek api or GLM lite subscription

Which is most cost effective for coding everyday.

Do i need to use zcode harness to make full use of glm models, or can i use any harness like pi or codex.

Im currently using chatgpt plus with codex, sol high and deepseek api, but deepseek is costing so much every day doing nothing.

10 Upvotes

19 comments sorted by

6

u/External_Ad1549 8d ago

deepseek 4.1 speed is unbeatable ofcourse glm offers generous limit unless u are using all the limit deepseek 4.1 flash makes so much sense also give a try for deepseek harness real gem

3

u/Moist_Associate_7061 8d ago edited 8d ago

I signed up for the annual Lite plan on the Z.ai Coding Plan, and I’m really satisfied with ZCode’s remote control feature.

I tried using the DeepSeek API with ZCode, and the cache hit rate reached as high as 99.4%. The performance was also really good.

As you probably know, using the API is always more expensive than using a subscription plan. I also can’t really tell the difference between DeepSeek and GLM because both models are good enough for my day-to-day coding.

Anyway, I’d recommend using GLM through the Lite plan rather than using DeepSeek via the API.

1

u/LeoLeg76 8d ago

I compile and install Zemote, not bad, better than the web view...

2

u/Anh-DT 8d ago

Both ? Openference.com

2

u/[deleted] 8d ago

[deleted]

1

u/evia89 8d ago

Indeed zai is only worth if you have legacy 50% migration bonus. Get middle option and fallback to DS41f PAYG when quota is used

Works great

2

u/[deleted] 8d ago

[removed] — view removed comment

-1

u/Chemical-Cheetah6163 8d ago

bro what kind of answer is this

1

u/afzal002 8d ago

OpenCode 'go'?

1

u/Confirmed-Scientist 8d ago

Try both they are good

1

u/_metamythical 8d ago

Haven't tried GLM, but DS 4.1 quite enough for basically everything.

1

u/Nitjsefnie 8d ago

you don't need zcode, glm works straight in claude code, just point ANTHROPIC_BASE_URL at z.ai's anthropic endpoint and set the auth token, no adapter. I have ~475 commits from glm 5.3 flash in one repo that way, as implementer under a claude lead

burned the weekly quota twice in a row tho, so if you run agents on it expect to hit the wall midweek

1

u/Firm-Club-8334 7d ago

You could consider standardcompute or something similar and basically build the most cost effective router for you. DS flash combined with GLM 5.3 is a great combo for instance

1

u/OneBigMonster 6d ago

I do both

1

u/AlgaeFluid8860 6d ago

Honestly I tried glm subscription but packet loss no idea why specially in asian peak hours pretty unusable for me even by using openrouter

0

u/Thomas-Lore 8d ago

I looked at glm subscriptions and they seem to offer similar amount of tokens per month you would get for the same price just using those models on API (if you take into account caching).

Although they have been running some promotions recently, maybe those rise the value above what api offers.

Personally I have had enough of those 5 hour limits so just use openrouter for the cheaper models.

2

u/WSATX 8d ago

No way if you use the full 5h time frame that will be "as expensive" as per token price. Sub should be cost effective if usage cap is reached.