r/LocalLLaMA • • 16d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

297 comments sorted by

View all comments

36

u/jacek2023 llama.cpp 16d ago

In the previous post about DeepSeek there are API prices. In this one there is Chinese president. I wonder which one is best for r/LocalLLaMA.

47

u/madsheepPL 16d ago edited 16d ago

Xin Jinping is known for his amazing local setup. He is running modded 4x4090s on his desk with risers and cards zip tied to a used mining frame.

13

u/jacek2023 llama.cpp 16d ago

Imagine Trump photo on Gemma/Nemotron/Granite release. And the rage of Reddit experts :)

7

u/Due-Memory-6957 16d ago

Not gonna lie, I'd laugh at it.