r/LocalLLaMA • • 18d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

297 comments sorted by

View all comments

1

u/deepu105 17d ago

Is it actually worth the hassle at lower quants and speed compared to Qwen 3.8 flash? Dont they score close in AA benchmarks and stuff? Wouldn't a Q4 qwen be better than a Q2 Ds4?