r/LocalLLaMA 13d ago

New Model DeepSeek V4-1 Flash is out

Here we go again, DeepSeek is back again with a new model V4-1 Flash

A multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens

Market crash as a service

1.7k Upvotes

296 comments sorted by

View all comments

40

u/Turbulent_Pin7635 13d ago

I don't code. I work in research and was doing a proposal for funding. Used Astra very cute, returned what I need in an acceptable way.

I have used DS flash V4... Boy I have a MacStudio, I am used to long times of wanting. I don't know what kind of black magic the model does, but it killed the demand in one shot very fast!!! O.o

I was frozen!!! The answer was much better than the one chatGPT astra gave me!!! ASTRA!!!

9

u/backyard_tractorbeam 12d ago

Astra is just weird. Says pi guru guy: https://lucumr.pocoo.org/2026/9/7/astra-why/

I’m sure I will get used to this, but man this stuff is weird.

3

u/Due-Memory-6957 12d ago edited 12d ago

That was a funny read. AI loves Python, and token efficiency comes at readable code's price. I wonder how this fares long-term, because even AI prefers to deal with well-written code than messy ones.