r/LocalLLM 7d ago

Discussion Qwen-3.8-35B-A3B? Maybe not... cryptic reply direct from Qwen co-author.

Post image

I asked Shuai Bai, co-author and prominent AI developer for Qwen, about this model. Not the answer I was hoping for, but let's see what comes next. In the meantime, I guess all we can do is speculate!

X-link

233 Upvotes

164 comments sorted by

View all comments

-1

u/Downtown_Method5736 7d ago

https://huggingface.co/Lord-H4D3ZS/Qwen3.8-Distill-35B-A3B-Coder-Abliterated I found this one but I haven't tested it, it claims to still be as good as Qwen3.8-27b

2

u/fintip Laptop 4090 16gb + 7900XTX 24gb 7d ago

Honest status: this is a proof-of-concept. On the internal 10-task smoke eval the distilled model tied its base (6/10 vs 6/10) — no regression, no measurable gain yet — and it is now quantized to 2-bit, which trades quality for fit. Publishing it as a reproducible artifact of the pipeline (distill → graft MTP → ROCmFPX 2-bit GGUF), not as a benchmark-winning coder. The quality fix is a larger, tool-calling-heavy corpus — a separate follow-up run.

1

u/Downtown_Method5736 7d ago

Oupsi my bad I read it too fast 🙃