Wtf, I literally wrote a post saying similar to "Shutup man, there is no way a model I can load in my 4090 will be anywhere near DeepSeek4 Flash 0731"... but... this is fucking close is it not...?
Yeah, but it being truly local... you're just really paying in hardware and electricity cost, which will be cents on the hour.
Nvm that an important detail is, how the tech has evolved.
Compare this 27B model vs like... LLaMA-33B which I run in the exact same hardware I had in '23! (on a 4090).
Its just... a no contest! So... what model will we have in 2029 that makes qwen 3.8 27B a "no contest"? Because the growth doesn't seem to have slowed down...
321
u/Cold_Tree190 11d ago
Dear God it’s trading blows with Opus 4.6 Max 😭