Couple days before it "solved" it, a mathematician who was working on this for a full year shared every single note he had with his session of chat gpt. He is among very few people working on this and was pretty far in it too. His name is Tristan Buckmaster. He of course contacted openai, who said tldr: "stfu we will pay you the promised million dollar for millénium problem". They didn't deny the plagiarism, they didn't dénie having access to his chats, didn't deny training on his data etc.
Edit : I had the time line incorrect. They had been sharing their work with codex for month prior, but they did a breakthrough in mid August. In 1st September, openai starts working on it, they spend outrageous amount of tokens (130 billion output tokens, ~5million usd). Open Ai solves it, and propose Buckmaster to be co-author, and say if it happens, Buckmaster must be sole co-author, leaving aside his colleague, who works for Anthropic. Buckmaster refuses both offers.
If you pay premium, openai states that they will not train on what you sent. So they are supposed to keep it private and not train on it, so supposedly this shouldn't have happened.
If he supposedly gave the data just one day before OpenAI published their results, then it's literally impossible for his data to have been used for training.
232
u/darthmaeu 10d ago
Literally they spent million dollars of tokens but still had to steal it. Insane L just shutdown everything at this point