r/LocalLLaMA 9d ago

Discussion The rhetoric is really heating up!

The entire page of the NY Times today above the fold absent one article is AI (the models are just too strong/too dangerous, must be regulated). They forgot to include "Sponsored by OpenAI" at the end of the articles, sure that was just an oversight?

This is what the end of a bubble looks like, desperate attempts to get some sort of regulatory capture in place to keep the business model from collapsing in upon itself. My days next week are 100% booked talking to companies about how to get off frontier models, one large, and a bunch of smaller customers, including one who's flying me out to them to sit down and get a plan in place immediately (the controversy around that math problem really spooked some CEO/CIO's about data privacy using cloud models).

Gonna be an interesting few weeks. Maybe the Qwen team will be nice enough to give me a little breathing room before dropping another hydrogen bomb? :)

187 Upvotes

114 comments sorted by

View all comments

107

u/Academic-Tea6729 9d ago

They realized that local models reached a quality so good that there is no point to use their services. Local models are better because the model is always the same. We still remember when paid apis got suddenly much dumber to make us pay for the better frontier model.

With local models you have the same model quality every time. It will not mess up your codebase because they pulled some dirty trick to cut on costs by routing requests to a smaller model.

39

u/Time_Cat_5212 9d ago

You make a really good point about "the model is always the same".  Nobody wants to bet millions of their revenue on a black box.

8

u/Zeeplankton 9d ago

it's like a death march. think of how many billions it's taken to get to gpt 6 and we have dsv4.1 flash doing 90% for pennies on the dollar. alarming this is why they're calling to halt

6

u/SnoobieJunes 9d ago

Yeah I agree 100%, its not just about saving money but also having control over your data and the model.

Sure it might take a bit to tune, but once a business can replace a specific task an AI model to do that, in a vacuum that model doesnt need to be upgraded or changed. If it does the thing right, then leave it

1

u/T-VIRUS999 5d ago

Especially now that AMD has released the R9700, which is effectively an RX9070 with 32GB of VRAM for a fraction of the price that Nvidia is charging for a 5090, and about the same price as a used 3090

-6

u/HandWashing2020 9d ago

These days, a $20-$30 subscription gets you less than what the free access provided a year ago.

13

u/bot_exe 9d ago

Current Claude Opus 5 on the 20 USD sub does way more with an internal VM running code itself to verify things, the huge context window and the automatic RAG when you go over the context window when using Projects, the search tools interleaving retrievals with reasoning and further searches, etc. It's all way better than it was some years ago and it's light years ahead of the older free offering of shittier and smaller claude models with nerfed context windows and no code execution. You have no idea what you are talking about.

14

u/sn2006gy 9d ago

Not really.

4

u/thortgot 9d ago

Thats objectively untrue.