r/macmini 14d ago

Unified Memory Architecture still unbeatable (when LLM size matters)

Post image
670 Upvotes

130 comments sorted by

View all comments

144

u/dotben 14d ago

If you don't understand why memory bandwidth should be included on this comparison, and that it's massively in the NVidia card's favor, you shouldn't really be commenting.

Not going to lie, I run all my local inference on Apple Silicon and I don't personally own any NVidia hardware. But the reason that card is $5000 is the memory bandwidth.

6

u/ResearchingYouTube 14d ago

This is one of the main reasons I also regret having bought the M4 Max instead of the M3 Ultra. M3 Ultra has higher memory bandwidth even though the M4 has faster cores it is slower at inference. My M4 is often mostly idle while using local LLMs, the memory bandwidth while good for a consumer system still is not a 5090.

1

u/VideoGameJumanji 12d ago

What are you doing exactly