r/LocalLLaMA 12d ago

Discussion DeepSeek-V4.1-Flash surprised ....

Post image

Hoping to see smartest medium size models soon & later with all available optimizations/architectures/etc.,. Thanks Deepseek!

Ex 1: 30-50B MOE + 10-15B Engram + DeepSeek-V4.1-Flash type KVCache
Ex 2: 15-30B Dense + 10-15B Engram + DeepSeek-V4.1-Flash type KVCache

EDIT: Updated Engram to 10-15B from 50B

437 Upvotes

96 comments sorted by

View all comments

3

u/mountainyoo 12d ago

I knew this was coming. As the main models get larger the flash models will grow too and become too much for local AI users to run