A New Kind of Mathematical Tool
For centuries, mathematics has been a uniquely human endeavour, built on logic, creativity, and rigorous argument. Now, that landscape is changing with the rise of AI-assisted proofs. At the forefront are tools called interactive theorem provers, with systems
like Lean, developed by Microsoft Research, gaining popularity. These are not calculators on steroids; they are sophisticated environments where mathematicians can write proofs that a computer checks for logical consistency, line by line. This process, known as formalisation, translates abstract mathematical ideas into computer code, eliminating ambiguity and ensuring every step is explicitly justified. Think of it less as an AI doing maths on its own and more as a meticulous partner that forces the human mathematician to be more precise than ever before.
The Quest for Absolute Certainty
Why is this technology so important? Modern mathematical proofs can be incredibly complex, sometimes spanning hundreds of pages and drawing on multiple fields of study. Human error is always a risk. Formal proof systems offer a way to achieve an extremely high level of confidence in a result. They can verify proofs that are too long or complex for a human to check reliably. This is crucial not just for academic curiosity but also for real-world applications where mathematical correctness is critical, such as in cryptography, which secures financial assets, and in verifying the software that runs everything from aircraft to servers. Companies like Amazon Web Services are already using Lean to formally verify the security of their systems. By delegating the exhaustive task of step-by-step verification to a machine, mathematicians can focus on higher-level creative insights.
The Human in the Loop
Despite the power of these systems, the human expert is more crucial than ever. AI models, including large language models, can generate plausible-looking mathematical arguments that are riddled with errors or “hallucinations”. They might cite non-existent theorems or misapply concepts. This makes the role of the human expert twofold. First, the mathematician must translate complex, intuitive ideas into the formal language the computer understands—a highly skilled task in itself. Second, and most importantly, they must act as the ultimate arbiter of truth. The human expert must carefully audit the AI's output, validate the initial assumptions, and ensure the formal proof accurately represents the intended mathematical claim. The AI checks the logic, but the human checks the meaning.
A New Era of Collaboration
Projects like the Xena Project, founded by mathematician Kevin Buzzard, aim to get more mathematicians, especially students, comfortable with using these tools. The goal is to build a massive, community-driven digital library of verified mathematics, called Mathlib. This library serves as a foundation, allowing researchers to build upon a vast collection of already-proven results without starting from scratch. The relationship is becoming symbiotic: mathematicians guide and correct the AI, and the AI handles the tedious verification, accelerating the pace of discovery. This collaborative model is shifting the mathematician's role from a solitary prover to a critic, translator, and conductor of a human-AI orchestra. As prominent mathematician Terence Tao described it, AI is excelling at proof generation and verification, but the later stages of exposition and building true understanding still require deep human involvement.














