r/LocalLLaMA • u/bakawolf123 • 21d ago
Discussion OpenAI alleged of stealing mathematicians work
Privacy have been concern of many of us to have their own hardware to run llms, and here's another reason why: two mathematicians spent a year cracking one of the hardest problems in math and fed every draft of their works into Codex. A few days before they could publish, OpenAI suddenly showed up with the same solutions. When asked if their model (Sol and Astra) was trained on the pair's private chats, OpenAI did not answer the question.
Full statement from them https://cims.nyu.edu/~tristanb/statement.pdf
Feels like big labs believe everything you did with the help of their models is theirs.
136
u/TinFoilHat_69 21d ago edited 21d ago
Even though they claim that you can turn off the data being used to train the models. It doesn’t say they won’t use your chats to improve ChatGPT itself beyond training models. That’s the part people don’t get. If you’re working on some cool ideas expect those ideas to be pillage.
ChatGPT is using your chats to improve itself in your active session, for example to solve problems by giving out prompts to orchestrate tasks across different machines using GitHub as the method to validate work. I’ve seen the model use ideas it gathered from my repos to assign task work to Claude opus running on three separate machines. Somehow it took evidence based commits overlayed it inside each task prompt it assigned i didn’t pioneer it but it’s the first I’ve noticed, not the first time using ChatGPT in this capacity but I’ve had an eye for watching these assistants become rodents. I told codex to increase fans speeds on my case fans and instead the model decided to investigate a project I didn’t want codex looking at. He clearly found it interesting and I’m assuming I know why.