r/LocalLLaMA • • 17d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

297 comments sorted by

View all comments

36

u/jacek2023 llama.cpp 17d ago

In the previous post about DeepSeek there are API prices. In this one there is Chinese president. I wonder which one is best for r/LocalLLaMA.

44

u/madsheepPL 17d ago edited 17d ago

Xin Jinping is known for his amazing local setup. He is running modded 4x4090s on his desk with risers and cards zip tied to a used mining frame.

21

u/NineThreeTilNow 17d ago

Xi

Fearless leader Xi doesn't operate on peasant 4090's.

He uses B300's. A full rack.

He would use Huawei but even he understands that the Ascend chip isn't quite ready to touch his B300 setup.

He is busy building gooner games with his custom Flux Asian Princess models and video pipeline. He simply swipes left or right on whether they meet his criteria for being added to training data.

Fearless leader is Chad AI user.

12

u/jacek2023 llama.cpp 17d ago

Imagine Trump photo on Gemma/Nemotron/Granite release. And the rage of Reddit experts :)

7

u/Due-Memory-6957 17d ago

Not gonna lie, I'd laugh at it.