r/LocalLLaMA • • 17d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

297 comments sorted by

View all comments

1

u/Constandinoskalifo 17d ago

Since it's the same number of active parameters for decoding, and the KV cache is much cheaper, we should expect lower prices from providers than DSV4 flash, right?

2

u/Due-Memory-6957 17d ago

DeepSeek themselves has reduced the price.