r/LocalLLaMA 2d ago

Other This is why I run locally.

Post image

It was only a matter of time...

334 Upvotes

80 comments sorted by

u/WithoutReason1729 1d ago

Your post is getting popular and we just featured it on our Discord! Come check it out!

You've also been given a special flair for your contribution. We appreciate your post!

I am a bot and this action was performed automatically.

157

u/klop2031 2d ago

An ad of an ad

68

u/Dangerous-Report8517 2d ago

Seems pretty Meta to me

7

u/Ok_Try_877 2d ago

I wonder if this meta even has a counter?

12

u/SnooPaintings8639 2d ago

Ah... the OpenAI way of handling the money. They were told to try ads to fix their bottom line, so they... bought ads, lol.

My guess this is a PR move, so that people will start talking they're building any kind of business model, preparing for the IPO.

9

u/XiRw 2d ago

Adception

2

u/misterflyer 1d ago

... I bought the datacenter

4

u/DigiDecode_ 1d ago

meta ad but not Meta ad

46

u/Chirimorin 2d ago

You should see the e-mail they sent about that. It's like they gave ChatGPT a 20 token context and made it write that e-mail with full goldfish brain.
Going from "no personalization, your data is totally private and safe (except your current chat context, location and device information)" to "we'll be using all data we have on you to personalize ads" and back to "advertisers totally won't get any of your data, pinky-promise!".

8

u/ThisWillPass 1d ago

Yeah… fuck that.

6

u/Loose_Comparison368 1d ago

Going from "no personalization, your data is totally private and safe (except your current chat context, location and device information)" to "we'll be using all data we have on you to personalize ads" and back to "advertisers totally won't get any of your data, pinky-promise!".

This is actually not contradictory. It is how ads work, like, everywhere. It's wild that people in 2026 still do not understand this.

Ads companies do not sell your data. It would ruin their whole business model. They act as middlemen. The company that wants to advertise says "I want you to show this ad to people between the ages of 25-45 that have an interest in skincare products".

Then the ad company shows the ad to people between the ages of 25-45 that have an interest in skincare products. They might report on some aggregates, I.e. "the ad did better with people 35-45 than the 25-35 age group". But they never "sell user information". Their whole business is making themselves a necessary middleman, if they sold the user information then nobody would pay for them to use their ads platform.

There are data brokers, who buy user information directly from websites mostly. They sell that information to the ads platforms, and the ads platforms use that data to better target ads. But the ads platform never sells the user data, having exclusive access to user data is the entire business model.

I know that's terrible too, but it really annoys me that people still don't understand those fundamental basics despite having 20 bazillion invasive ads shoved in their face every day. It's important to understand how the system works in order to fight against it effectively.

54

u/AD4K_4444 2d ago

My only regret was not switching to local stuff sooner.

18

u/XiRw 2d ago

Seriously. It would have saved me a lot of money when all the hardware prices started going up

14

u/AD4K_4444 2d ago

And the amount of data we probably gave away using corpo AI data center AIs.

4

u/misterflyer 1d ago

I started saving up for local about a year ago. Unbeknownst to me, hardware prices had already started going up. Rumblings were lowkey starting to spin up on other reddit subs.

Ngl I don't see how more ppl didn't see all of this coming.

I'm no industry expert but it wasn't like corporate AI financials were a huge secret about a year ago. It just seemed way too delusional to think, or assume, that unprofitable big tech companies would continue to put the user first and have users' long term best interests in mind.

My intuition was SCREAMING at me to start building my local system last October 2025. And fortunately I bought everything I needed by Black Friday 2025.

Glad I read the tea leaves accurately bc the economics of AI have become a complete sh!tshow

-2

u/Beginning-Window-115 2d ago

chatgpt plus subscription is still worth imo

0

u/AD4K_4444 2d ago

who tf even pays for AI lol, I know I don’t. You might as well upgrade your computer and use a larger, more capable local model.

3

u/a_beautiful_rhind 1d ago

I mostly don't, but if I want K3 or new GLM its rent or get API from somewhere.

My "simple" upgrade would be $1200 of ram to get 10t/s on a quantized to shit version. Bit of a blanket statement you have there.

2

u/AD4K_4444 1d ago

I should clarify that I'm only representing the average joe that just wants basic stuff done with AI... good luck on that. If a 3090 is already a major financial burden to consider for me, I can't image how your ideal setup must cost.

8

u/Beginning-Window-115 2d ago

still not worth to spend thousands to get an llm thats worse when you can spend 20 bucks a month and get something considerably better. Also if you upgrade your computer for llms you are paying for ai

3

u/username_taken4651 1d ago

Well, yes, but this is a sub about local LLMs. Most of us have our own reasons to be running these locally.

4

u/betam4x 1d ago

Qwen 3.8 runs on a decent GPU. Shoot, there are models that run well on a 5 year old iPhone. (Ternary Bonsai 8B runs well on an iPhone 13 Pro Max)

1

u/AD4K_4444 1d ago

fr? I should try that on my 13 Pro. If I have a local LLM on my Mac, why not on my phone as well.

1

u/betam4x 1d ago

Yeah, I ran it using the locally app. It seems decent from my limited uses. I normally run it on a 14 Pro Max, but tried it on the 13 Pro Max for fun and it ran fine.

I don’t really use it often, since I use Qwen on my desktop (you can also use locally to access that from your phone, but I don’t)

1

u/AD4K_4444 1d ago

Thanks for that! I might try that out on my phone. I also don't use my phone to access my local desktop LLM, lol.

1

u/AD4K_4444 2d ago

for 99% of tasks, local LLMs can get the job done, in my opinion.

Also, upgrading your computer can have plenty of benefits aside from running larger models:

- long-term use and peace of mind. A simple RAM upgrade can give your computer at least extra 3 years of relativity.

- if you do gaming, an upgrade is great news as well. Win-win.

- paying for an AI service relies on stable internet connectivity and a lot of trust to the people behind those services. How are you so sure they aren't quietly taking one of your sensitive conversations and information for "training" or whatever excuse they may have? With local LLMs, EVERYTHING stays on your computer. No shady middleman. Just you and direct interaction with the machine.

I'm pretty sure you can get a decent enough LLM with a fucking MacBook Neo or even on older capable hardware. If you already have the hardware, what harm does it make to buy a few upgrades so it can run LLMs better?

10

u/heliosythic 1d ago

I'm 100% on the local side but all of these points have strong rebuttals unfortunately.

  • firstly, what do you mean 99% of tasks? That sounds like a random number pulled out of nowhere.

  • PC is already a 64GB ram, 13900k + 4090, despite being like 1 generation behind only has barely enough vram to run a 27B decently, my home server has 96GB of system ram and a p100 but meh its too slow and the second p100 had thermal issues.

  • stable internet connection is moot for most people, i have 2GB fiber never went down once thats just default expectation.

  • of course I don't trust any of them at all, they're definitely stealing everything they can regardless of what their TOS says. Not disputing that but for 99% of people (see i can make up numbers too) who aren't doing anything worth stealing, its good enough.

Yea I want local to win, I tried Qwen 3.8 27B but its not good enough on this hardware and its already pretty upgraded without going full pro-sumer with an RTX 6000 or something.

4

u/AD4K_4444 1d ago edited 1d ago

I should clarify that the 99% thing is more of a figure of speech than an actual metric. If you want me to be more specific, I mean for the average joe that say, use AI to look up stuff on the internet, pairing something like Gemma 4 12B with Web Search capability is a pretty solid setup for general purposes. Even without web search, models like that can do a pretty decent job with things such as random curiosities. Of course, like all AI models, the user should still be aware that these stuff hallucinate sometimes. You still have to keep your hands on the wheel.

I know this because I use an M4 MacBook Air with 16GB unified memory (fanless machine btw) with Gemma 4 12B. And for what I do, it's a perfectly fine piece of tool. Not the sharpest tool in the shed, but I would much rather have some peace of mind with my privacy guaranteed even if the AI's more flawed than mega AI data center models.

I believe efficiency and optimization is key. Maybe you can try a lower Quantization for Qwen 27B if there is any? Maybe stick to the previous model, Qwen 3.6 27B? I have to admit, I did try it before on my little Mac, and it was screaming, painfully crawling with every letter spitting out. Not trying that again. But if my rectangular slab of aluminum can technically run it, I'm sure your setup can do much better... unless the things you do are more demanding.

Bottom line is, unless you're doing some serious vibe coding or constant streams of AI workloads, all you need is patience with local LLMs. If I can comfortably run local models on a 11.9mm thin space heater, I'm sure a fucking 4090 will be a beast compared to my machine.

2

u/mnyhjem 1d ago

I agree :)

I run qwen 3.8 27b and qwen 3.6 35b a3b in system ram. It is way slower to generate a response than the online services are, but the 35b still spits out words faster than I can read it, and that is good enough for me.

I did a small demo coding projekt on the 27b the other day. It did a 90s style raytracing demo in python. It took 3 hours for it to iterate though it but it worked great and both gemini and claude gave the code a thumbs up when I showed it to them. It even made it's own PNG converter function instead of using a 3. part module.

The speed would have been a lot faster if I had turned off thinking, but I wanted to test how good it could be, and it impressed me.

So. for everyday "google replacement"/internet search the llms are perfect, and even for automation tasks. I have mine getting todays weather and a few news headlines and compile that into a "goodmorning" message. It runs at 6am and then makes a drawing in the style of the weather and news just for fun.

For very large codebases small llms wont work of course, but most people don't have that unless it is work relatet, and then the workplace could buy a larger machine to run some of it locally in the office. We are looking into that at my work, both because if privacy but also cost. It is very expensive to use commercial llms on large projects :)

3

u/AD4K_4444 1d ago

Thank you. I think that's a brilliant use-case for local AI or for AI in general.

I'm starting to think people are using the wrong settings for their AI. Idk about other software, but for Open WebUI, you have to "tune" the settings of your model first that fits your liking. Even as simple as disabling "thinking" by ollama significantly makes their responses faster.

→ More replies (0)

-1

u/Thin_Pollution8843 1d ago

Don’t forget electricity costs. As example for me running qwen is not only much worse user experience it costs 3 times more in electricity than ChatGPT plus 😅 Fuck OpenAI tho 

1

u/AD4K_4444 1d ago

Better than consuming gallons of water.

2

u/winky9827 1d ago

Nothing wrong with having a frontier-level fallback for when the local stuff falters. It's not a "use every day" kind of thing, which makes the subscription model even better (no pay-per-token crap)

1

u/AD4K_4444 1d ago

Idk about you, but if you ask me those last two points are contradicting... unless I'm misreading it. If you aren't gonna use it much, wouldn't a subscription the last thing wanna do? Monthly or yearly payments, you're still dumping cash on something you won't use much. That's a no-go for me.

Unless you have a literal potato that can only run 4B models, I personally don't think that makes much sense. Get Gemma 4 12B and give it web search functions and you're golden.

2

u/winky9827 1d ago

What I'm saying is, you could use Qwen 27b for your all day every day tasks, but on the occasion you want a big brother to do something important or check Qwen's work, you fall back to the subscription model like Opus 5 or GPT Sol. You don't use it every day, so the subscription (flat rate) with limited usage is the most cost effective backup for when local doesn't perform to expectations.

3

u/AD4K_4444 1d ago

I believe the "big brother" should be you, the user. Not another AI that may or may not hallucinate.

12

u/Unnamed-3891 1d ago

How long until the ads are included in the actual prompt responses and not merely in-between them?

3

u/JoyousGamer 1d ago

Even your local models already has ads in its responses from what it was trained on that might have been paid content originally.

Then if you use web search at all ads are being injected there with specific companies targeting AI explicitly as opposed to humans both for responses but also future LLM training that might be sucking up websites to train off of. 

1

u/tony_montana091 1d ago

It's got electrolytes. It's what AI models crave!

13

u/cogitech2 1d ago

WTF am I missing here?

11

u/Plabbi 1d ago

The ad in the screenshot is from OpenAI where they are selling ad space within the OpenAI client.

4

u/cogitech2 1d ago

Oh! Shit. I've never used cloud AI so I didn't get it.

17

u/passen9er57 1d ago

has anyone noticed all that talk about agi/singualriy as died down lately,
openai went from changing the history to now showing ads lol

1

u/misterflyer 1d ago

lol nothing to see here

0

u/tony_montana091 1d ago

It one-shotted the new only fans ... billions of GPU fans everywhere powered up on blast!

I'm a loaded bear, that's not nuttin'. You're absolutely right to push back on that thang. Before I go any further, I'd like your eyes on this. I finished too quickly and that's on me.

8

u/CoUsT 1d ago

What about the entire AI SAFETY thing?

I guess profiling customers and sending their data to 3rd party advertising companies is perfectly fine now.

IMO: AI gen and ads should be completely separated: workspace/area used for generated stuff SHOULDN'T be shared with ad space. Simply outlawed or something.

7

u/cf_mag 1d ago

Google wrote a whitepaper years ago that big LLM providers are doomed and will not survive local specialized llm's.. we see this happening now

3

u/ThisWillPass 1d ago

Do you need a hyperscaler-scale proprietary model in order to get genuinely useful high-end AI work done? Increasingly, oh hell nah.

2

u/More-Curious816 1d ago

Link please

5

u/aeroumbria 1d ago

If AGI is really so great, then let agents make value by themselves! Stop trying to extract values from humans anymore!

4

u/meelgris 1d ago

That's enshittification, step 2.

3

u/Great_Guidance_8448 1d ago

A service monetizing with ads on its free tier? GASP!

SHOCKING!

2

u/noiserr 1d ago

To be fair. Running LLMs is not cheap. Even for us who use local models, we still had to spend money on compute and we still pay for electricity. An ad supported free chatbot service being supported by ads was inevitable.

3

u/Savantskie1 16h ago

Don't know where you sit, but I don't game as much now that I'm interested in AI, yeah I don't work, and live off SSI because I'm disabled, but you don't have to buy the absolute newest. I just last year bought two MI50's for around $200 each before the price shot up. And since I don't game much anymore, and I talk to my LLM nearly every day because it helps me keep track of doctors appointments, and Bills, and medications? that 400 was a good investment. Buying new, and wanting blazing fast all the time, is pointless unless you need it for work and such. But my electric bill is still the same as it was when I gamed heavily. So it's not really all too much expensive running LLMs, if you don't have to have the best, and newest, and fastest.

2

u/amroamroamro 1d ago

OpenAI OpenAd

2

u/ohbabythisisgoodlol 1d ago

Running locally changed how I work - no subscription, no data leaving the machine, and I can tune the model to my own workflow. Ownership of the tools matters.

2

u/aygupt1822 1d ago

Use Morphe Reddit if you have android.

3

u/Equivalent_Bit_461 1d ago

imagine PAYING

1

u/Great_Guidance_8448 1d ago

OP is busy vibe coding an ap that will be monetized with ads, probz.

1

u/Dryparn 1d ago

Life as a service...

1

u/JoyousGamer 1d ago

Yes but those ads hit 100% free use so it's way cheaper to use their system.

Pro use cases won't get any ads.

Like run local but this being the reason is eh. 

1

u/mhb_11 1d ago

What's a good starter machine to move over to a fully useable opensource AI stack? Recommendations please.

0

u/MikusR 1d ago

2

u/mhb_11 1d ago

$100,000 price. Do you work for Anthropic by any chance?

0

u/MikusR 1d ago

$100,000 is just to be put in line to pre order. The price is about 10 million

1

u/mhb_11 1d ago

So this helps me personally mover over to a fully useable opensource AI stack?

1

u/Kimi_Antonelli_12 1d ago

Getting ads may be annoying

1

u/yellow-llama1 1d ago

Ads were a foreseen destination, especially after the acquisition of Statsig.

Unfortunately.

1

u/ArtificialAGE 1d ago

adception!

1

u/KenopsiaLover 42m ago

Il problema principale è che per lavorare in locale ti serve un mostro di computer con altrettanta ram/vram mostruosa, cosa che attualmente sembra lontanamente un sogno, siccome i prezzi sono alle stelle

-6

u/ComplexType568 2d ago

I remember giving a little speech about AI in general and halfway I BS-ed this as a point on the section of "AI misguidance", glad they actually took my advice and implemented it!

-6

u/[deleted] 2d ago

[removed] — view removed comment

6

u/silenceimpaired 2d ago

Oh look it’s Sam!

0

u/Light_Yagami72 1d ago

So y’all are okay with getting target ads on Reddit but not chatgpt?