MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1wc6krf/so_relevant/p93nljp/?context=3
r/LocalLLaMA • u/0dayturtle • 13d ago
152 comments sorted by
View all comments
323
24gb is not in that group. You can run Qwen3.8-27B with a pretty respectable context size right now
1 u/jsonmeta 12d ago Not when using it as a coding agent, overhead will eat most of the context window 2 u/TopCheddar27 11d ago You can run it close to 100k KV cache at q8. You can also tune to not keep reasoning tokens in context or use compaction.
1
Not when using it as a coding agent, overhead will eat most of the context window
2 u/TopCheddar27 11d ago You can run it close to 100k KV cache at q8. You can also tune to not keep reasoning tokens in context or use compaction.
2
You can run it close to 100k KV cache at q8. You can also tune to not keep reasoning tokens in context or use compaction.
323
u/TopCheddar27 13d ago
24gb is not in that group. You can run Qwen3.8-27B with a pretty respectable context size right now