What Did The AI Actually Do?
On October 6, 2026, OpenAI published a massive collection of mathematical findings generated by one of its internal, unreleased AI models. The company released 722 manuscripts, organized into 372 groups of related results, on the code-hosting platform
GitHub. This wasn't a typical academic paper release; it was a flood of information intended to push the boundaries of several fields, including number theory and algebraic geometry. The AI attempted around 4,000 different problems, with each successful result requiring, on average, the equivalent of three hours of computing power from a professional ChatGPT subscription. The results include progress on some of the most famous unsolved challenges in mathematics, such as the Riemann hypothesis and the Unique Games Conjecture.
A New Kind of Scientific Partner
This event marks a significant shift from AI being a tool that simply follows instructions to one that can explore complex, abstract concepts and generate novel hypotheses. Unlike previous breakthroughs, which often involved a huge swarm of AI agents and massive computing costs, most of these new results came from a single prompt given to a single AI agent. This suggests the AI isn't just using brute force but is developing a more nuanced form of reasoning. To make the findings verifiable, many of the proofs were released with formalizations in 'Lean', a programming language that allows a computer to check the logical steps of a mathematical argument. This is crucial, as the sheer volume of AI-generated work could easily overwhelm the capacity of human mathematicians to review it all manually.
Why 'Expert Checking' Is The Real Story
Despite the AI's power, the 'expert checking' part of the headline is the most critical element. The results are not presented as infallible truths. OpenAI itself acknowledges that proofs lacking formal verification may contain errors and has committed to addressing any issues identified by the community. This positions the AI as a powerful but unproven collaborator. Human mathematicians are still essential to provide the final sign-off. Their role is to verify the correctness of the proofs, interpret the significance of the findings, and understand the new ideas or techniques the AI might have uncovered. The process is less about an AI replacing mathematicians and more about creating a powerful human-machine partnership to accelerate discovery.
A Community Divided
The release has been met with a mix of excitement and apprehension within the mathematics community. Some, like distinguished professor Alex Kontorovich, expressed astonishment, noting that if a human had produced one of the results related to the Riemann hypothesis, it would be an 'instant Fields Medal, no questions asked'. However, others are more cautious. The speed and scale of the AI's output have led some to worry that mathematics could become a field where results are produced faster than they can be understood. In an open letter, 25 Fields Medal winners warned about a 'severe misalignment' between the goals of the AI industry and the practice of mathematics, which thrives on deep understanding, not just answers. This tension highlights a growing debate about the role of AI in intellectual and creative fields.
















