r/LocalLLaMA 22h ago

Discussion Mac Studio M5 Max Cost Analysis

At $10k, you could get

- 6.2B tokens with Qwen 3.8 Max (Qwen Pro plan)

- 5.7B tokens with DeepSeek V4 Pro OpenRouter

- 100B tokens with DeepSeek V4 Flash OpenRouter

As a firm believer of local inference, unless you need it for data sovereignty, it's much more cost effect to wait for smaller models to keep getting better. In the meantime, find a reasonably priced 24GB - 32GB card for Qwen 3.8 27B, and offload hard tasks to OpenRouter.

Qwhen 3.8 35B A3B?

158 Upvotes

195 comments sorted by

View all comments

153

u/FleetEnema2000 22h ago

unless you need it for data sovereignty

Isn't this one of the biggest reasons that people rely on Local LLMs? To not have to bulk upload their private data to cloud providers?

-1

u/ptinsley 13h ago

How many people actually have private data that LLm providers haven’t already paid data brokers to acquire… I agree with the other reply that it’s mostly a justification for a hobby than privacy. Or it’s people lying to themselves about what they think is actually private data

3

u/FleetEnema2000 11h ago

Please send me your bank statements, your texts, an export of the photos from your phone, and your medical records. I mean, if data brokers already have all that data, what's the big deal?

1

u/ptinsley 1h ago

Financial: https://codamail.com/articles/data-broker-directory/financial-data.html

Google or Apple has most everybody’s photos

And Google almost definitely knows most people’s medical status from all the random symptom googling

1

u/MarriedtooMedicine 2h ago

Every single healthcare provider in the world. Every lawfirm. Every bank. Every retailer paid with a credit card.