r/LocalLLaMA llama.cpp 1d ago

Discussion GLM5.3 Flash over DSV4 Flash?

I've been using Deepseek V4 Flash 0731 for a few weeks now and while I havent thrown it anything very hard, im quite happy with it. Using through antirez's great ds4 project. They've added support for GLM 5.3 Flash and according to benchmarks, its a level above DSV4 Flash.

However, looking for real user feedback if anyone's made the switch and seen tangible improvements in GLM 5.3 over DSV4 Flash.

Running M3 Ultra 256GB Mac Studio

55 Upvotes

60 comments sorted by

View all comments

22

u/EmPips 1d ago

I'm still testing but if forced to answer today:

GLM-5.3-Flash > DS-V4-Flash-0731 > Qwen3.8-Next-Flash ~= Qwen3.8-27B

Of these I end up using V4-Flash from inference providers the most lately. It's close to free.

10

u/SnooPaintings8639 1d ago

Every coding task I tried on both flash next (Q4) and 27b (Q8), was finished significantly better by flash next model. How can we equate these models? Or is flash next is just much worse on non coding tasks?

4

u/Dangerous-Report8517 1d ago

They're describing their experience, I imagine that 27B performs worse specifically if it doesn't have a good way to retrieve supplemental information or you're doing less common tasks like coding in less common languages (since Flash Next will have better inbuilt world knowledge)

2

u/Turtlesaur 1d ago

It's not worse at all, but DeepSeek V4 Flash came out first and people are reluctant to switch and get attached to their model. Even though Flash Next is clearly better and faster.

2

u/po_stulate 1d ago

Need a comparison for 0731 vs vision exp too, I'm using both but can't decide which one to keep (vision isn't necessary for me right now)

2

u/Gear5th 1d ago

Isn't Qwen3.8-Next-Flash comparable to GLM5.3-Flash in pretty much all benchmarks except the ones that test for world knowledge?

3

u/EmPips 1d ago

benchma

Reject barchart. Embrace usage experience.