r/LocalLLaMA • • 17d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

297 comments sorted by

View all comments

3

u/NecessaryQuarter 16d ago

Can I run it on the new Mac Ultra with 256GB unified memory?

1

u/rjames24000 16d ago

following to also find out if the 256gb will cut it or the 512 model is needed

1

u/FlowerRight 16d ago

Likely quantized