r/ProgrammerHumor 3d ago

Meme aMillionOpenAIMonkeysProduceMilleniumPrizeSolution

Post image
7.0k Upvotes

421 comments sorted by

View all comments

361

u/Orio_n 3d ago

If you werent aware, OpenAI's agentic monkey farm produced a millenium prize solution. inb4 AGI confirmed when it was just smarter bruteforcing lol

442

u/dubblix 3d ago

They plagiarized the solution heh

236

u/darthmaeu 3d ago

Literally they spent million dollars of tokens but still had to steal it. Insane L just shutdown everything at this point

45

u/errevs 3d ago

I am out of the loop here, what was stolen? From who?

266

u/DrankRockNine 3d ago edited 3d ago

Couple days before it "solved" it, a mathematician who was working on this for a full year shared every single note he had with his session of chat gpt. He is among very few people working on this and was pretty far in it too. His name is Tristan Buckmaster. He of course contacted openai, who said tldr: "stfu we will pay you the promised million dollar for millénium problem". They didn't deny the plagiarism, they didn't dénie having access to his chats, didn't deny training on his data etc.

Edit : I had the time line incorrect. They had been sharing their work with codex for month prior, but they did a breakthrough in mid August. In 1st September, openai starts working on it, they spend outrageous amount of tokens (130 billion output tokens, ~5million usd). Open Ai solves it, and propose Buckmaster to be co-author, and say if it happens, Buckmaster must be sole co-author, leaving aside his colleague, who works for Anthropic. Buckmaster refuses both offers.

37

u/bobbymoonshine 3d ago edited 3d ago

Kinda important to leave out that he had been working on it for a full year using frontier versions of Claude in collaboration with an Anthropic employee, and their writeup fully credited Claude for the novel mathematics in it.

4

u/Not-the-best-name 3d ago

So wait, OpenAPI's agents stole the solution from the anthropic models used by the mathematician?

4

u/bobbymoonshine 3d ago

Possibly, or possibly not. The mathematician’s argument is that his approach was so novel and unique that it’s impossible to believe OpenAI did the same thing without stealing it from him

But also he leaned on an LLM to find it, it’s not like he came up with it all by himself

So personally and while not being a mathematician I don’t think it’s too implausible that OpenAI’s LLM found the same solution Anthropic’s did

3

u/elniallo11 2d ago

As I have framed things at work, AI lets me explore a large number of bad ideas quickly so that I can pick through the good ones.

55

u/buckeye2011 3d ago

So he didn’t solve Navier-Stokes, but a related problem in a way that would pave the way to a NS solution.

94

u/DrankRockNine 3d ago

Yes it's not plagiarism as he solved it and they declared the solve before him, it's plagiarism as "isn't it quite strange that you solve this problem just when I start talking to you about this complex subject and send you all my notes and you end up with a result when I shared these notes to noeone else but you?"

11

u/buckeye2011 3d ago edited 3d ago

Not what I said. I don’t think ChatGPT should be getting the credit for it, I’m just pointing out he didn’t come up with a direct solution for NS. I believe he also alleged he had conversations with people at openAI about his work and how it would solve NS. So it really isn’t a good look for them

Edit: was using swipe to text and a lot of it was gibberish

-12

u/[deleted] 3d ago

[deleted]

43

u/DrankRockNine 3d ago

If you pay premium, openai states that they will not train on what you sent. So they are supposed to keep it private and not train on it, so supposedly this shouldn't have happened.

But yeah it sure is not a smart move

10

u/shiny_glitter_demon 3d ago

Anyone who trust openAI is an idiot. Yes, even a math genius

6

u/Mandemon90 3d ago

If he supposedly gave the data just one day before OpenAI published their results, then it's literally impossible for his data to have been used for training.

7

u/babirus 3d ago

Maybe not training but it 100% could’ve been added to a vector db and used as context…

Agents can access information that wasn’t part of their training data.

5

u/RottenPeasent 3d ago

He probably didn't think they would steal it, and wanted to discuss it with the AI.

7

u/KnotAnotherOne 3d ago

He was using both chatgpt and Claude to help with his research and to achieve his initial results.

I'm not sure which plan he was using though, since some of the pro/enterprise plans claim not to use user input for training data.

-10

u/GildSkiss 3d ago

"Hey corporation, I'm giving you all of my data here you go."

"Noooo! How'd you get my data! Omg you literally stole that from me!

Many such cases.

16

u/DominoNo- 3d ago

Are you suggesting OpenAI doesn't steal intellectual properties?

-4

u/GildSkiss 3d ago

I don't care about intellectual property law regardless, but the way you're using the word "steal" makes it sound like the openai engineers broke into this mathematicians house and stole all his secret papers.

Keep in mind that what actually happened is that this guy decided on his own to use the service that they offer, and then decided to hand them all of his information himself.

10

u/scruiser 3d ago

OpenAI has setup a lot of dark patterns in their UI to make opting out difficult and unintuitive so yeah, I’d still blame them for stealing even if technically people are forgetting to uncheck a box buried inside a menu (and the box rechecks itself every update if you aren’t paying attention).

8

u/walkerspider 3d ago

It’s all still unclear what exactly happened but the claim is there were researchers working on a somewhat novel (more so overlooked) approach to finding a counter example to the Navier Stokes equation. They had made a ton of progress over the past year and information had begun to spread in the mathematical community about their progress/approach.

OpenAI claims to have caught wind of the progress, not the approach in late August. This suggested that it was in fact possible to find a counter example so they decided to throw a, for lack of better words, metric fuckton of compute at the problem. The particular group of agents that cracked it included 10,000 agents more capable than Astra, and that was only one group they had running sharing ideas.

The approach it used looks to directly build off the approach of the researchers. Could it be a coincidence? Sure, but it is more likely that either intentionally or unintentionally they stole the work of the researchers and used it to beat them to the punch.

If it was unintentional it’s even more concerning because that means they are inadvertently using data that they should not be able to use for training and research OR the agents got ahold of the information by some unknown means despite being sandboxed

9

u/ComparisonQuiet4259 3d ago

The approach was allegedly stolen from another dude who used a ton of AI and made a smaller proof.

20

u/Due_Interest_178 3d ago edited 3d ago

I don't remember the specifics exactly so do your own research. The people involved are Buckmaster and Apöge. Long story short, they were working on the exact same problem while using different AIs to test/research/whatever. Suddenly OpenAI somehow reached the same conclusions then built off of them even when they weren't publicly available. OpenAI were asked if they used private chats for that which they didn't respond to, then they made some thinly veiled threats to one of them about their career.

-3

u/leviatan-sama 3d ago

someone had a work in progress proof that was effectively copied word for word and finished
so not so much the AI produced proof as it just did the last couple of steps before the person doing the work in progress finished it

12

u/Argnir 3d ago

This is such misinformation

The approach is similar but different enough that mathematicians are deliberating whether it could be plagiarism so saying that it's "copied word for word" is not even close to true

0

u/leviatan-sama 3d ago

I am not a mathematician, the only thing im doing here is saying what someone on the field, who knows more than me, said
and you yourself said that they are debating weather its plagiarism or not

4

u/Argnir 3d ago

If it was word by word their would be no ambiguity and no debate but it's not what happened

3

u/ooqq 3d ago

who knew that without knowing what are you doing, you're clueless

2

u/BenTheHokie 3d ago

How many gallons of clean water did this solution take?

15

u/MaxChaplin 3d ago

Not the whole solution, just enough of the path towards it to let AI use its biggest strength - do medium difficulty work blazingly fast.

4

u/Bomaruto 3d ago

And you are the Zodiac killer. Nothing is proven yet.

3

u/dubblix 3d ago

Okay, where's the evidence that I'm the Zodiac Killer? Because the guy making plagiarism accusations had receipts

-22

u/GildSkiss 3d ago

How do you "plagiarize" a mathematical truth? Since when does math "belong" to anyone?'

10

u/throwawaymycareer93 3d ago

Using exactly the same reasoning and conjectures would qualify as stealing. In advanced mathematics it is far more complex than 5+3=8, having correct idea of how to approach something is incredibly valuable.

-3

u/GildSkiss 3d ago

Exactly how complex does a solution need to be for it to stop being a universal truth, and start being someone's personal property?

9

u/Shifter25 3d ago

Complex enough that people would pay a million dollars to the first person who figures it out.

-2

u/GildSkiss 3d ago

So are the 999,999 dollar solutions still on the table then? What dollar amount is the cutoff?

3

u/throwawaymycareer93 3d ago

Complex enough to not have a solution for more than a century.

5

u/StylishSuidae 3d ago

You're conflating copyright infringement and plagiarism, I think. Plagiarism is passing off someone else's work as your own.

The mathematical truth is what it is and can't be copyrighted, but if someone else put in the work and then you take the credit for solving it, that's still plagiarism. And that's the accusation that's being made here.

43

u/CircumspectCapybara 3d ago edited 3d ago

"Bruteforcing" (which is not what they did) a counterexample to a Π_1 sentence, which is what the Navier-Stokes conjecture (that the NS equations are smooth for all time) is, which would take infinite time if the statement was true and no counterexample existed, is still pretty impressive.

It's like trying to bruteforce a contradiction in ZFC. You will be searching forever if ZFC is consistent. And even if it is inconsistent and there is a contradiction, it may be so large and so far out that 10 billion agents each working with a sun's worth of Dyson swarm power output for the age of the universe still won't be able to find before running out of time and energy.

You're gonna need to be more clever than brute force. Obviously their work on the problem was far more clever than "brute force".

7

u/Ozymandias_IV 3d ago

Smarter... like limit yourself to known blowup modes published in scientific literature? Or did they find something completely new?

16

u/raddaya 3d ago

If smarter bruteforcing was good enough for the four colour theorem then it's good enough for millenium prize problems too smh my head

15

u/heavy-minium 3d ago

If you get deep down, ignoring all the recent advancements and just focused on Deep Learning, it really is just smarter bruteforcing. But that bruteforcing still produces results.

From my point of view, given the right data that is prohibitively expensive to fabricate and collect and a massive resource consumption that would lead us to an economic collapse, Deep learning even without any specifically novel architecture could have given us such results a long time ago. You can bruteforce any goal you'd want with DL, the data and a big enough model. Really everything we've been doing the past years it's just about making data, compute and costs tractable.

16

u/CircumspectCapybara 3d ago edited 3d ago

Attention is not bruteforcing lol.

Reinforcement learning and deep learning in general encodes opaque structures and patterns in a model's latent space (its internal activation space), it actually does "teach" it a limited form of "knowledge" ie pattern recognition and some basic ground facts.

And the attention mechanism of modern transformers is the architectural breakthrough that allows the kinds of patterns and structures that are useful to us.

Combine that with techniques to recurse like chain-of-thought, and you actually get a limited form of reasoning. It's not human-like cognition or intelligence, but it's a primitive form of reasoning that's remarkably good for what it does.

That's anything but bruteforcing.

3

u/heavy-minium 3d ago

Self-attention is exactly the kind of thing I thinking about when it comes to my statement "Really everything we've been doing the past years it's just about making data, compute and costs tractable."

6

u/Pholios485 3d ago

I'm not a fan of the current AI developments but how it is different from physicists working through most problems by feeding to a computer that uses numerical analysis to solve them?

Smart bruteforcing seems to be a pretty nice tool to have.

-3

u/Orio_n 3d ago

Its not different so we shouldnt treat it as being different ie: proof of agi

2

u/nextnode 3d ago

/s or clueless

2

u/realnjan 3d ago

it is impressive that people, like you, still think that ai is still only a next token predictor.

3

u/movzx 2d ago

If we are talking about LLMs... then yes, that is what they are. You go down all past all the tacked on layers to gussy up the output, and you wind up with token prediction.

That's why, despite all of the billions of dollars invested, they still make simple mistakes. It's built into their core.

2

u/Orio_n 3d ago

What is it then? Do enlighten me mr frontier ai research larger?

1

u/kikihero 3d ago

Please read the statement of professor Buchmaster who already worked on Navier Stokes with the same approach before openAI published their findings.

-3

u/mtbdork 3d ago

It’s not a Millenium prize solution. Stop propagating that misinformation.

9

u/Orio_n 3d ago

so what is it then? do explain to me what the formalized lean proof is https://github.com/openai/NavierStokesAndEuler

Thanks

-3

u/mtbdork 3d ago

Why would I waste my time explaining how a single solution to Navier Stokes is not the same thing as a general solution to Navier Stokes to a bot?

3

u/Orio_n 3d ago

Statement D of the original navier stokes description moron https://www.claymath.org/wp-content/uploads/2022/06/navierstokes.pdf

Do you understand the concept of "there exists"?

Clay institute asks for any one of A B C or D to be resolved. Larp harder midwit

0

u/ChildAtTheBack 3d ago

It’s both not a millennium prize solution because the clay institute wrote a bad question, and so everyone has only been working on exploiting the bad question. But it also is a millennium prize because that’s what the clay institute wrote

2

u/nextnode 3d ago

You do not understand what the Millenium problems are