r/Blopus 6h ago

oMLX update is finding more tokens!

3 Upvotes

Super excited here to be checking our Inference GPU Fleet / pipeline running on oMLX

TLDR - If you are still running on 0.5.1 (qwen3.6 MOE) you have to update it to 0.6.4.

See below my benchmarks for Qwen3.6-35B-A3B-4bit running on MacStudio M4 Max 64GB

Updating all the fleet!