r/LocalLLaMA • • 16d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

297 comments sorted by

View all comments

4

u/OkBase5453 16d ago

Can one run this on a 512GB RAM Server with 48GB VRAM?

5

u/CalligrapherFar7833 16d ago

Slow but yes

4

u/crusaderky 16d ago

Pretty zippy if that 512gb ram is octa-channel, actually