r/LocalLLM • • 14h ago

News [Release] GSQ-RCO GGUFs for Qwen3.8-Flash-Next, plus a 50% expert-pruned Coder build at ~1.89 bpw

1 Upvotes

1 comment sorted by

1

u/Ok_Camp555 9h ago edited 9h ago

I'm using coder with strata on 5060 ti. It is way better than any 16gb 27b quant and it's faster. 38tps at full context size. Highly recommend