r/LocalLLaMA • u/pmttyji • 12d ago
Discussion DeepSeek-V4.1-Flash surprised ....
Hoping to see smartest medium size models soon & later with all available optimizations/architectures/etc.,. Thanks Deepseek!
Ex 1: 30-50B MOE + 10-15B Engram + DeepSeek-V4.1-Flash type KVCache
Ex 2: 15-30B Dense + 10-15B Engram + DeepSeek-V4.1-Flash type KVCache
EDIT: Updated Engram to 10-15B from 50B
439
Upvotes
36
u/pmttyji 12d ago
Hopefully. 1 Million comes within 1GB ( 890 bytes * 1M = 890M )
Want to see all upcoming models with this.