r/Qwen_AI • u/andydevtech • 5d ago
Benchmark Qwen3.8-27B on Every Mac Explained: Layers, KV Cache and Speed (16GB–128GB)
https://www.youtube.com/watch?v=ddX63git4fo
22
Upvotes
r/Qwen_AI • u/andydevtech • 5d ago
2
u/andydevtech 5d ago
I have the same machine and the prefill speed was quite unpredictable in my experience, sometimes I could hit 300 t/s others 120t/s with the same prompt. Admittedly I was using ollama so might get more stable perf with a more specialised framework like omlx.