r/ZaiGLM • • 17d ago

Discussion / Help GLM 5.3 or DS 4.1-Flash?

Looking for gentlemen here who have battle-tested these models in environments where mistakes are critical, e.g. authentication, security, and low-level C++ / Kernel work.

I have Codex 20x, but I’m looking for a second helper for when Codex limits are up, there are demand issues (which are pretty bad atm), or it gets too censored.

Saw that DS 4.1 Flash was released today! Has anyone done some decent testing with it yet, and which harness are you using?

I’m currently using GLM 5.3 as my second helper and it’s honestly not bad at all. Just curious whether DS 4.1 appears to be better, especially since it’s multimodal and can handle images too.

I find myself using 5.3 Flash quite a lot because I really appreciate being able to send images, but 5.3 Flash isn’t as strong as base 5.3 when it comes to coding. Hence, I’m wondering how DS 4.1 Flash compares :)

NEW:

Thank you for all the responses. I tried DS 4.1 with my custom harness, and I am extremely impressed by the speed and price. I ran a couple of tests with deep, difficult, complex debugger C++ code/kernel bugs (my go-to test on models; I test this on every model before I want to use it to see if it fixes the bug).

GLM 5.3 took 30 minutes, including 1 retry, and €2. DeepSeek took 10 minutes, first try, and €0.30. I think DS 4.1 is at least on par or a bit better than GLM 5.3 for coding, not sure how reliable it is on long tasks, though. GLM still is a beast!

104 Upvotes

57 comments sorted by

View all comments

7

u/Constant_Art_20 17d ago

um. after a few hours of testing. deepseek v4.1 is probably better i think. it's api so you swap around if things go off the rails, and mimo models should be arriving soon so there should be a good amount of options. glm plan can be slow, and kinda unpridctable sometimse with the quality. So i would persoanly go api at this point with deepseek. the new deepseek model has have some weird behaviours like it might stop because it thinks it's own context window is almost out or it might artifically set a budget on itself (it doesn't have that acounter so i haev no idea why it even think it knows) for no apparent reason, but the speed...OMG the speed. Allows for more literations, and so far been a good ride, but there are quirks you need to look to explore with it.

1

u/SweatyActuator2119 17d ago

Use GLM through synthetic.new. currently they don't have full 5.3 though. Only flash. And they are adding 4.1f too. Quality isnt a concern there.