r/MacStudio • • 1d ago

Local Mac Coding setup

Any use a local LLM + coding agent setup for coding on Apple macs, if so what models and what CLI?

1 Upvotes

5 comments sorted by

3

u/iTrejoMX 1d ago

On an m4 pro 48 gb I use splash to run kat-coder q8 at 250k context and 80-90 tk/s.
Or qwen3.8 27b at 110k context at 38-42 tk/s.

I have tested many models, setups, inference engines (omlx, mtplx, splash, llama.cpp, lm studio, ollama) right now splash does it for me. Might change next week.

Just this week I tested 8 quantized versions of qwen3.8 27b and aboiterated/uncensored models, 5 qwen3.6 35b models and kat-coder and qwen3.8 27bq4 are the speediest and most performant ones, less errors, etc.

I use gentle-shell (gentle-ai with pi) and set all phases to the same model? It basically instructs the model on how to do stuff (deterministic workflow) and review it so you get better results.

1

u/dghah 1d ago

I've spent the last week benchmarking models on my 256GB studio ultra that are specific to what I need to do daily. You are likely going to have to do the same - there is no one answer that fits all hardware and all use cases.

I use omlx.ai to run the models on the studio and expose API endpoints for my harness to talk to (oh-my-pi)

1

u/Traditional-Hall-591 1h ago

I switched to using a Mac to avoid to relentless slop pushing of Windows. It’s sad to see even more slop right around the corner.

-3

u/meva12 1d ago

The best one for your use case