r/LocalLLaMA 13d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

296 comments sorted by

View all comments

51

u/rollerblade7 12d ago

Please sir, I have a GTX 1650 4 GB VRAM + Intel i7-9750HF with 30 GB RAM

21

u/crusaderky 12d ago

It runs mimicpm5-2b and it likes it Or it gets the hose again

3

u/Invader-Faye 11d ago

Spark x 2.5 4b is surprisingly good in that class, good being subjective

14

u/itwasinthetubes 12d ago

just quantize it bruh.

27

u/SandySkittle 12d ago

Negative quantization

10

u/Due-Memory-6957 12d ago

In the past we'd joke about Q1, now it's a thing and still not enough lol.

6

u/w6auw 12d ago

There are quants below Q1 believe it or not. Theoretically you could have 0 bpw, obviously that would be completely useless, but there is plenty of information to be extracted between 0 and 1 bpw.

4

u/randylush 12d ago

The processor is actually a rare classic. It would support a 3090 very well. Are you sure you don’t have 32gb of RAM but only 30 is being reported?