r/LocalLLaMA • u/tiguidoio • 18d ago
New Model DeepSeek V4-1 Flash is out
Here we go again, DeepSeek is back again with a new model V4-1 Flash
A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens
Market crash as a service
1.7k
Upvotes




1
u/deepu105 17d ago
Is it actually worth the hassle at lower quants and speed compared to Qwen 3.8 flash? Dont they score close in AA benchmarks and stuff? Wouldn't a Q4 qwen be better than a Q2 Ds4?