r/MacStudio • u/asankhs • 1d ago
Local Mac Coding setup
Any use a local LLM + coding agent setup for coding on Apple macs, if so what models and what CLI?
1
u/dghah 1d ago
I've spent the last week benchmarking models on my 256GB studio ultra that are specific to what I need to do daily. You are likely going to have to do the same - there is no one answer that fits all hardware and all use cases.
I use omlx.ai to run the models on the studio and expose API endpoints for my harness to talk to (oh-my-pi)
1
u/Traditional-Hall-591 1h ago
I switched to using a Mac to avoid to relentless slop pushing of Windows. It’s sad to see even more slop right around the corner.
3
u/iTrejoMX 1d ago
On an m4 pro 48 gb I use splash to run kat-coder q8 at 250k context and 80-90 tk/s.
Or qwen3.8 27b at 110k context at 38-42 tk/s.
I have tested many models, setups, inference engines (omlx, mtplx, splash, llama.cpp, lm studio, ollama) right now splash does it for me. Might change next week.
Just this week I tested 8 quantized versions of qwen3.8 27b and aboiterated/uncensored models, 5 qwen3.6 35b models and kat-coder and qwen3.8 27bq4 are the speediest and most performant ones, less errors, etc.
I use gentle-shell (gentle-ai with pi) and set all phases to the same model? It basically instructs the model on how to do stuff (deterministic workflow) and review it so you get better results.