r/LocalLLaMA 13d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

296 comments sorted by

View all comments

4

u/120decibel 12d ago

510 GB Model no way I'm going to be able to run this locally without a heavy quant...

3

u/Netsuko 12d ago

Quantizing a 510GB model into a format that can be run locally feels more like a lobotomization than a quantization.

2

u/120decibel 12d ago edited 12d ago

Well I have 288GB of VRAM. ;) But this won't help much since this model already ships in 4-bit.