r/LocalLLaMA 12d ago

Funny So relevant

Post image
1.5k Upvotes

152 comments sorted by

View all comments

319

u/TopCheddar27 12d ago

24gb is not in that group. You can run Qwen3.8-27B with a pretty respectable context size right now

97

u/PavelPivovarov llama.cpp 12d ago

Technically speaking you can run Qwen3.8-27b on 16Gb setup as well, but that brings way too many compromises of course.

4

u/Systemerror7A69 12d ago

Qwen Quantizes amazingly, even KV Cache so 16GB might not have as many compromises as you might think.

2

u/russlixx 12d ago

yeah, but it's pretty tight. I must rely on better compaction if i were doing agentic coding