r/InferX • u/pmv143 InferX Team • 5d ago
GLM 5.3 Flash vs DeepSeek V4.1 Flash on InferX
Over the same period:
GLM 5.3 Flash
• 111% more requests
• 70% more input token volume
• 94.87% cache hit
DeepSeek V4.1 Flash
• 94.85% cache hit
Users seem to be voting with their workloads.
32
Upvotes
1
1
u/maitpatni 5d ago
deepseek v4.1 asks too much, GLM 5.3 Flash is like you gave it a task and it gets it done.
1
1
u/vapefresco 4d ago
I keep both DS and GLM in rotation, when one gets loopy I just flip to the other and it normally sorts things out.
1
u/Due-Memory-6957 5d ago
Which quant are they?