OpenAI Claims GPT-5 Solves Century-Old Math Problem
OpenAI claims GPT-5 solved the century-old XYZ number theory problem, potentially advancing cryptography. Mathematicians remain skeptical due to opaque logic and lack of public source code for verifiโฆ
OpenAI announced on Tuesday that its AI system has produced a proof of the conjecture known as the โXYZ problem,โ a longโstanding puzzle in number theory that has stumped mathematicians for over a century. The company released a short statement on its website, claiming the proof was generated by its latest language model, GPTโ5, and that the model had independently derived the result using a novel algorithmic approach. The announcement came during a week of heightened interest in AIโassisted research, following DeepMindโs success with AlphaFold and a recent paper from researchers at MIT showing that neural networks could find new patterns in algebraic geometry.
The XYZ problem, formally stated in 1953, has been a touchstone for the field because it links prime numbers to complex analysis. Solving it would close a chapter on a central question about the distribution of primes and could unlock new cryptographic techniques. The timing of OpenAIโs claim is no accident: the company has been investing heavily in mathematics research, and the release of GPTโ5 coincides with a broader push to demonstrate AIโs capacity for โhighโlevel reasoning.โ Experts note that AI systems are now capable of manipulating symbolic expressions and exploring vast search spaces far beyond human reach, making them promising tools for tackling deep conjectures. However, the mathematics community has long been cautious, insisting that any claimed proof must undergo rigorous peer review before being accepted.
Mathematicians have reacted with a mix of intrigue and skepticism. Dr. Elena Kovรกcs of the University of Cambridge said the proof โcontains many novel steps, but the logic is not yet transparent enough for us to confirm its validity.โ Others point out that the proofโs source code was not made public, raising questions about reproducibility. A group of researchers on Twitter has begun to trace the proofโs lineage, finding similarities to a preprint published last year by a small team of independent researchers. OpenAI has denied any plagiarism, stating that the model was trained on publicly available data and that the proof was generated in real time during the experiment. The company has offered to release the modelโs internal logs to a panel of independent reviewers, but it has not yet provided the full proof for external scrutiny.
The next step will be a formal vetting process. If independent mathematicians can verify the proof, it could be submitted to a leading journal such as the Annals of Mathematics. Even if the proof is flawed, the episode will likely accelerate discussions about AIโs role in scientific discovery and the standards required for machineโgenerated claims. The controversy also highlights the tension between openโsource AI research and proprietary development, as OpenAIโs approach to sharing its findings remains cautious. Whether the XYZ problem is finally solved will hinge on the communityโs willingness to engage with the evidence and on the transparency of the tools that produced it.
Read Full Story at Scientific American โ


