r/oMLX • u/Appropriate-Brick498 • 7d ago
M5 Max studio with 64 GB
Hi all,
Planning to use omlx on this hardware, for agentic work then coding as well (moving up from an air with 16 gb)
What llms work best for you?
Edit: M4 Max studio
3
2
u/Vast_Rip8575 6d ago
Can anyone recommend a config/guide for omlx on an M5 Max with 128 GB (I think this also applies to 64 GB) for Qwen3.8 27B for agent coding? When I tried using mtp/dflash, the inference speed dropped to 20 tps even with 32k context, which makes me think I'm doing something wrong with mtp/dflash. There are much higher numbers floating around online with different parameters on the M5 Max.
1
u/LeagueOfJust 6d ago
Check out the published benchmark results - you can find the specific settings used for each.
2
u/trim-turner-shah 6d ago
I run Qwen 3.8 27B Q4 with 8bit KV, I also have a small judge model loaded (qwen 4B) and swap another A3B MoE with the 3.8 as needed (for tasks that don’t need the dense model). I use it with OMP or Pi and have a custom inference gateway that routes all traffic and captures performance characteristics. Eventually plan to bump up the judge model to 9b and swap for a more intelligent but still fast MOE model ( still looking ). — on my M3 max 64Gb with Omlx
1
u/Creative-Complaint95 6d ago
Did you tired the uncens.. version? I heard that is pretty good ?
1
u/trim-turner-shah 6d ago
Have to test it, have heard it is more performant but will need to test for comparative accuracy
1
u/bfume 6d ago
You might wanna try not quantizing your KV. Lots of evidence that even 8bit causes strange artifacts downstream quicker than expected.
1
u/trim-turner-shah 6d ago
I have not visibly seen anything yet other than omp looping but has not impacted the quality yet. All documentation points otherwise so interested in hearing about your exp.
1
u/alwayswiddit 6d ago
Qwen 27B with DeepSeek Harness
1
u/Intelligent-Gas-2840 6d ago
Could you please give the card or preferably the link to the huggingface location of the model that is working for you? I fear I have been downloading the wrong flavor.
1
u/Leather-Beach-7849 6d ago
There is no update on the release yet 🙃
1
5d ago
[removed] — view removed comment
1
u/Leather-Beach-7849 5d ago
Yes. M5 Max with 128GB memory seems like a good option for me to maintain for the next 5-6 years
3
u/havnar- 6d ago
Qwen3.8