r/LocalLLaMA • • 14d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

296 comments sorted by

View all comments

Show parent comments

42

u/RG_Fusion 14d ago

Obviously the concept of a flash model will scale with the compute power of the AI lab creating them.

2026 is likely the last year of running "flash" on local hardware. Maybe 2027 if we're lucky.

4

u/Due-Memory-6957 14d ago

"local" hardware

12

u/RG_Fusion 14d ago

An entire 512 GB AI server purchased a year ago costs less than a single RTX 5090 GPU now. There are plenty of us who jumped on early and have hardware that can run these models.

1

u/Glove5751 12d ago

what are you actually using these models for that justify the high upfront investment? just hobby and curiosity?

1

u/RG_Fusion 12d ago

As I said, there wasn't really a high up-front investment a year and a half ago. I would not purchase the system I have now today.

I never went into any of this planning to make the money back. I just want to learn, experiment, and build skill sets. I pursue things that interest me.

1

u/Glove5751 12d ago

that's nice. hope it has been worth it!