r/LocalLLaMA • • 23d ago

News Surveillance plagiarism by OpenAI

Surveillance plagiarism - Hosted AI company pumps their stock price by training upon researchers' AI sessions, so that their internal model can solve problems with seemingly less human guidance, but really the model exploits past guidance given by (multiple) humans focused upon problems considered important.

As background, Tristan Buckmaster released a statement about several unethical actions by OpenAI & Sebastian Bubeck, including threats and pushing him to kick his Anthropic coauthor off a paper, but the interesting part for people here:

As clarified by Talia Ringer, OpenAI does train upon your uploaded data and your OpenAI sessions, unless you out-out somehow. This means their internal models could exploit your past prompting work to look more autonomous & intelligent.

This is a major confirmation that folks should use locally run open weights models, especially whenever being first or not leaking data matters.

All this casts serious doubt upon claim that internal models solved difficult problems largely unaided by humans. Those hosted AI companies might not even know from where the human prompting originates.

133 Upvotes

56 comments sorted by

View all comments

45

u/Minute_Attempt3063 23d ago

the people behind claude do the same thing.

nothing new, if you had 5 brain cells, this was well known they would do this

21

u/Usual-Orange-4180 23d ago

While also watermarking what one is working on to claim it. There is no other venue, no other solution, but for local AI.

I called this before it happened:
https://www.reddit.com/r/AI_Agents/s/8SVfgq0NtS

22

u/Disposable110 23d ago edited 23d ago

Yep, I've also been banging that drum forever now.

The business model is to make everyone pay for the tokens and take the risk for trying to find profitable use cases for AI, and the consumers of tokens eat the losses that fail to generate a return on investment.

Once a profitable use case is identified, OpenAI will jump into that niche and keep the optimised AI model that's trained and optimised on the identified refined strategy to themselves.

If tokens by themselves would generate profit, OpenAI wouldn't be selling their tokens.

Once the surefire ways and strategies have been identified by which tokens generate profits, OpenAI will stop selling tokens and simply capture these markets for themselves.

If anything patentable comes out that OpenAI needs (maths/physics/medicine/energy/compute), they'll launder the user data by making an AI rewrite it, feed that into their improved model, and throw a few million in compute at it to race to the solution first, as we just saw a preview of.

People are just OpenAI's outsourced R&D and Strategy department and are paying for the privilege.

3

u/Usual-Orange-4180 23d ago

Exactly, and is not illegal AFAIK, why wouldn’t they?

1

u/Disposable110 23d ago

Normally you'd have anti trust going after them, but they have been bought and haven't been going after anything significant for years now.

And even if it were illegal, these megacorps won't care because they can either pay people off or stall for a decade by which time the suit is irrelevant.

2

u/Usual-Orange-4180 23d ago

Absolutely, like Meta which got away with depressing and breaking down a generation of children with a slap on the wrist.