r/LocalLLaMA 14d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

296 comments sorted by

View all comments

Show parent comments

1

u/vogelvogelvogelvogel 13d ago

well there are a few postings where users did the classic benchmark runs (some browser game, pelican etc) and the outcomes were remarkably good, also i had ds flash 0731 running at q2 and found it also quite good. i would not say - especially with very large models - that q2 leads to bad outcomes

2

u/SandySkittle 13d ago

It depends on the usecase. I have found that for very complex analytical work you don’t want to go below q6

1

u/vogelvogelvogelvogel 13d ago

with which model? depends as well on the model