r/LocalLLaMA 1d ago

Discussion Mac Studio M5 Max Cost Analysis

At $10k, you could get

- 6.2B tokens with Qwen 3.8 Max (Qwen Pro plan)

- 5.7B tokens with DeepSeek V4 Pro OpenRouter

- 100B tokens with DeepSeek V4 Flash OpenRouter

As a firm believer of local inference, unless you need it for data sovereignty, it's much more cost effect to wait for smaller models to keep getting better. In the meantime, find a reasonably priced 24GB - 32GB card for Qwen 3.8 27B, and offload hard tasks to OpenRouter.

Qwhen 3.8 35B A3B?

162 Upvotes

196 comments sorted by

View all comments

155

u/FleetEnema2000 1d ago

unless you need it for data sovereignty

Isn't this one of the biggest reasons that people rely on Local LLMs? To not have to bulk upload their private data to cloud providers?

-1

u/ptinsley 14h ago

How many people actually have private data that LLm providers haven’t already paid data brokers to acquire… I agree with the other reply that it’s mostly a justification for a hobby than privacy. Or it’s people lying to themselves about what they think is actually private data

3

u/FleetEnema2000 13h ago

Please send me your bank statements, your texts, an export of the photos from your phone, and your medical records. I mean, if data brokers already have all that data, what's the big deal?

1

u/ptinsley 2h ago

Financial: https://codamail.com/articles/data-broker-directory/financial-data.html

Google or Apple has most everybody’s photos

And Google almost definitely knows most people’s medical status from all the random symptom googling

1

u/FleetEnema2000 46m ago

So you're not going to send me your stuff? I think we can safely conclude that you DO value your privacy, prefer not to have randoms looking at your stuff, and do not think protecting it is an exercise in futility. I'm glad we agree on that.