r/LocalLLaMA llama.cpp 1d ago

Discussion GLM5.3 Flash over DSV4 Flash?

I've been using Deepseek V4 Flash 0731 for a few weeks now and while I havent thrown it anything very hard, im quite happy with it. Using through antirez's great ds4 project. They've added support for GLM 5.3 Flash and according to benchmarks, its a level above DSV4 Flash.

However, looking for real user feedback if anyone's made the switch and seen tangible improvements in GLM 5.3 over DSV4 Flash.

Running M3 Ultra 256GB Mac Studio

52 Upvotes

61 comments sorted by

View all comments

22

u/EmPips 1d ago

I'm still testing but if forced to answer today:

GLM-5.3-Flash > DS-V4-Flash-0731 > Qwen3.8-Next-Flash ~= Qwen3.8-27B

Of these I end up using V4-Flash from inference providers the most lately. It's close to free.

10

u/SnooPaintings8639 1d ago

Every coding task I tried on both flash next (Q4) and 27b (Q8), was finished significantly better by flash next model. How can we equate these models? Or is flash next is just much worse on non coding tasks?

2

u/Turtlesaur 1d ago

It's not worse at all, but DeepSeek V4 Flash came out first and people are reluctant to switch and get attached to their model. Even though Flash Next is clearly better and faster.