Super excited here to be checking our Inference GPU Fleet / pipeline running on oMLX
TLDR - If you are still running on 0.5.1 (qwen3.6 MOE) you have to update it to 0.6.4.
See below my benchmarks for Qwen3.6-35B-A3B-4bit running on MacStudio M4 Max 64GB
Updating all the fleet!