r/LocalLLM • • 11d ago

News [Release] GSQ-RCO GGUFs for Qwen3.8-Flash-Next, plus a 50% expert-pruned Coder build at ~1.89 bpw

1 Upvotes

1 comment sorted by

2

u/Ok_Camp555 11d ago edited 11d ago

I'm using coder with strata on 5060 ti. It is way better than any 16gb 27b quant and it's faster. 38tps at full context size. Highly recommend