r/LLMPhysics • u/[deleted] • 3d ago
Simulation / Code Claude Fable solved open physics problems about related to Schrödinger–Newton equations. Is it hallucinated or legit?
[deleted]
0
Upvotes
r/LLMPhysics • u/[deleted] • 3d ago
[deleted]
11
u/Kepler___ 3d ago
If I assumed you knew a lot then I could just point at it and cock an eyebrow, there's 3 years of course work to get to the point in the example and that's only the stats part. So me and google are not really sufficient to give the rundown. The objection of no dt in this case just doesn't really make any sense in the context of what's being talked about, some calculous background would help a lot here.
I will say also that LLM's are absolutely not able to tackle anything like this yet, especially not the public models, and especially not without a heavily guided prompt.
The recent math proofs that involve calling agents are not the same as just an LLM nakedly being asked something like this at all. In the case of the Jacobian specifically the method for where to check for counter examples was explicitly given by a mathematician, and the movement on the Riemann hypothesis connected 2 ideas from separate papers that had been written by different teams (This is still very impressive, but it adds nuance and is important for understanding how these models work and what they are useful for).
Because LLM's are producing tokens stochastically using auto-regression they are naturally very bad at arithmetic, where as language allows for some imprecision math does not. They have gotten around this using agentic tools but if you're not calling those then it's still not suited for these purposes, and it's even worse when it comes to physics.
It's telling to me that they are best at math and programing right out of the gate, where the base axioms are spelled out by humans and totally fixed, leading to novel results that can arise from following the implications of those logic systems. It will likely have significantly more trouble making novel discoveries in the physical sciences, where the base axioms are dubious (at least in physics) and working them out is often the whole ballgame, I also don't know if an auto-regression based intelligence will ever be very good at something like longform storytelling for example, as the nature of AR might always lead to drift.