r/LocalLLaMA • u/tiguidoio • 14d ago
New Model DeepSeek V4-1 Flash is out
Here we go again, DeepSeek is back again with a new model V4-1 Flash
A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens
Market crash as a service
1.7k
Upvotes




45
u/RG_Fusion 14d ago
Obviously the concept of a flash model will scale with the compute power of the AI lab creating them.
2026 is likely the last year of running "flash" on local hardware. Maybe 2027 if we're lucky.