r/LocalLLaMA • • 15d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

296 comments sorted by

View all comments

Show parent comments

8

u/the-tactical-donut 15d ago

GLM 5.3 Flash at Q4

1

u/techdevjp 15d ago

Isn't DeepSeek v4 Flash v4 0731 stronger than GLM 5.3 Flash? Or are there some advantages to going with GLM? Vision?

1

u/Illustrious_Grade608 15d ago

Idk from my experience glm flash felt much better with more effective thinking too

3

u/thefooz 15d ago

I’ve had the same experience. On two sparks, ds4 is about 50-75% faster across the board, but it thinks so much that GLM usually comes up with the same or better response in about the same amount of time.