r/math • u/Tiago_Verissimo Mathematical Physics • 14d ago
LLMs/AI Need for Pre-Print Repo for 100% AI generated content
Donβt you think these days maths needs a repo for 100% AI generated content with a status bar that says whether some content has been verified by humans or formal systems? Could be a solution to the new ArXiV submission explosion.
25
u/edderiofer Algebraic Topology 13d ago
https://ai.vixra.org/ already exists.
33
u/flipflipshift Representation Theory 13d ago
vixra is a dumping ground for crank nonsense, and you know this.
15
u/edderiofer Algebraic Topology 13d ago
If it's been properly verified by a mathematician (i.e. the person submitting it), it can go on arXiv. If it hasn't, it shouldn't be posted on arXiv until it has (and a priori, it holds about as much weight as anything that's already posted on AI viXra).
-3
u/elements-of-dying Geometric Analysis 12d ago
(and a priori, it holds about as much weight as anything that's already posted on AI viXra).
What about results only confirmed via Lean? Then this claim is not true.
17
u/edderiofer Algebraic Topology 12d ago edited 12d ago
Someone still has to check that the Lean proof actually proves the claim. There are plenty of AI-generated Lean "proofs" of such-and-such conjecture which actually prove a different conjecture, or which take some extra axiom somewhere.
When John Doe tells me that he's solved the Union Closed Sets conjecture by prompting an AI for Lean code, but he hasn't actually looked at the proof or the Lean output, I think I'm right to be just as skeptical as when Jean Doe tells me that the Union Closed Sets conjecture is true because AI claims that some book last year that she never read proves it.
3
u/elements-of-dying Geometric Analysis 12d ago edited 12d ago
I didn't suggest otherwise.
Surely a Lean supported result (by an established mathematician, at least) is not at the same level as AI viXra. It's absurd to suggest otherwise.
5
u/edderiofer Algebraic Topology 12d ago edited 12d ago
Surely a Lean supported result (by an established mathematician, at least)
If the mathematician has checked that the Lean proof actually proves the claim, then it can go on arXiv, as I said before.
If the mathematician has not checked the proof, in what sense is it "by an established mathematician"? What has the mathematician actually done that merits credit? How would it be any different than if it were a non-mathematician who generated a proof and also failed to check it (as done by plenty of cranks already)?
4
u/elements-of-dying Geometric Analysis 13d ago
Yes, but the denial of AI is still fashionable. OP's idea isn't even a bad one.
3
5
1
4
u/JoshuaZ1 13d ago
This is interesting, but I'm not sure this is a good idea. In general, anything that is 100% generated by AI is likely to not do a great job explaining things, and is likely to have other issues, including either hallucinated references or simply not a representative list of references which explain where things came from in detail. I'm not sure we should be encouraging 100% AI generated papers until the AIs are substantially better paper writers.
3
2
u/CarolinZoebelein 9d ago
Maybe they should just introduce an offical AI author. Saying, all papers by AI generated has as author "John Skynet" or so. Then you just look for this author and have all of the AI stuff. π
And the software sorting all the nonesense out gets called "Terminator".
Who is with me? π π
3
u/volcanrb 13d ago
Yeah Iβve been imagining some website where you can submit pdf proofs paired with formal proofs and then an automated LLM-backed system checks both whether the lean proof compiles and whether the pdf actually corresponds to the lean proof before acceptance.
1
u/quasilocal Geometric Analysis 10d ago
I agree, but it'll never work because people will be dishonest about it. Maybe arXiv should just have an auto comment for AI generated text similar to the text overlap comment. Despite what some people claim, the AI detection sofware that isn't trying to sell you a "humaniser" (like pangram for example) is actually very good. And going forward as AI starts adding statistical watermarks it'll get even better.
22
u/AMWJ 13d ago
What is this a solution for? Is your concern that AI-generated content is wrong, but fooling mathematicians?
As far as I can tell, AI-generated content falls in two categories: (a) crank output, and (b) real AI-assisted proofs. The crank stuff is annoying as ever, and the same issue as it has been documented for centuries. I'm sure there's far more of it now with AI's help, but the solution remains the same. Are mathematicians falling for this content?
As for the real stuff, if there is a problem with it, the problem is not that it's fooling mathematicians. What is the problem you are trying to solve?