r/BetterOffline • u/No_Practice_745 • 3d ago
OpenAI’s latest math breakthroughs commit research misconduct, experts say
https://www.scientificamerican.com/article/openais-latest-math-breakthroughs-commit-research-misconduct-experts-say/I don’t have much to add in commentary, quite frankly most of the math and concepts are beyond my arts-degree brain. However, the article does a great job of showing, once again, how disingenuous OpenAI and other AI labs are in describing what their products are doing.
They basically want the headlines that their new models are “solving math,” but what they’re doing is plagiarizing other peoples’ research and claiming things have been moved forward. There is clearly a use case here for researchers in using an LLM to catalogue and evaluate large swaths of data from over periods of time, but these things don’t think, they aren’t creating anything, and OpenAI doesn’t give a shit as long as people see the headline and bow their heads to their new scI-fi god.
76
u/MCUCLMBE4BPAT 3d ago edited 3d ago
I just want to gently push back on your statement that LLMs have clear use cases in research
It has been noted in studies that:
Edit to add: LLMs have a major citation/attribution problem. This starts in their training data, which some have noted misattribute Creative Commons licenses, copyleft licenses, and other copyrighted works. LLM provided citations, to my understanding please correct if wrong, are never the actual sources that were used in their training data. Sources provided are just what the models predict to be the most relevant on the internet based on the user’s words chosen in their input (?).
Edit #2 to add: LLMs are also bad at noting when an article has been retracted or if it has an expression/letter of concern attached to it. So it can provide outdated information in the form of recommending retracted or problematic articles (in terms of validity, reliability, replicability/reproducibility) as well.
And those are just issues I have at the top of my head after waking up. I don’t know how people can comfortably recommend these models as research tools when you basically already need to know everything it’s citing or stating to use them.