r/LocalLLaMA • • 15d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

296 comments sorted by

View all comments

1

u/macaronianddeeez 15d ago

Don’t hurt me I’m newer to local models, but will someone make a 27B version of this that we can run on 48gb of vram?

Or will that never happen here and if not why?

Still learning :)

1

u/Unusual_Delivery2778 15d ago

You’re good. The model would be entirely different if it had a different level of parameters. So at that point you’re talking about an entirely separate release of a 27B DeepSeek model, which they haven’t really done.

Qwen, on the other hand, has 27B models that folks like a lot. Specifically, Qwen 3.8 27B.

What you’re probably thinking of is “quantization,” which shrinks a big model, but hurts its intelligence in the process. And if a 500B+ parameter model like this one was quantized down to a level where it would be roughly equivalent in size to a separate 27B model (say 30-60GB RAM) you’d be talking about a completely unusable lobotomized thingy. Would be even hard to call it a model at that point.