AI vs. Fermat's Theorem: Revolution or Formal Mirage?
Anthropic succeeds in formalizing Fermat's Theorem in 11 days using agents, reopening the debate on the true nature of mathematical reasoning.
September 7, 2026 · 4 min read

TL;DR: Anthropic has used Claude agents to formalize Fermat's Theorem in 11 days, a milestone that automates large-scale mathematical verification. While it is an unprecedented technical achievement, experts question whether AI truly adds creative value to mathematical knowledge or if it merely accelerates mechanical work.
The automation of mathematical truth
The recent feat achieved by the orchestration of Claude agents, developed by Anthropic, in formalizing Fermat's Theorem in just eleven days using the Lean language, is not merely a technical achievement; it is an epistemological rupture. Historically, formalization—the process of translating natural mathematical language into computer-verifiable code—required years of highly specialized human labor. The fact that 13 million lines of Lean code and over 30,000 intermediate theorems were validated in less than two weeks marks the end of an era in pure research. If we compare this milestone with Andrew Wiles' original proof in 1994, which took seven years of voluntary isolation to consolidate decades of work, the difference in scale is abysmal. AI has not 'discovered' the theorem, but it has reduced the cost of verification from 'years of human life' to 'cloud computing cost,' altering the economics of mathematical proof.
Why is this a turning point?
Formalization acts as an audit of reality. In modern mathematics, an error in a chain of reasoning can invalidate decades of derived theorems. The use of Lean (a proof assistant based on type theory) allows computers to act as infallible arbiters. The paradigm shift here is subtle but profound: we are moving from 'proof by consensus' (where experts manually review the work of others) to 'proof by compilation.' This leap is comparable to the transition from manual accounting to ERP systems in the 20th century; like ledgers, mathematical logic is being migrated to an environment where human error is technically impossible. For the mathematician, this means that value no longer resides in the technical execution of the proof, but in the formulation of the original conjecture and the logical architecture that guides the agents.
Old-school skepticism
The reaction of the lead researcher, who had a five-year grant for this task, encapsulates the tension between 'truth' and 'understanding.' His statement: 'The result tells us little about the essence of mathematics', resonates with the historical criticism of calculating machines since the invention of Pascal's calculator. There is a justified fear that mechanical efficiency might obscure human intuition. While AI operates through the brute force of coordinated agents, human mathematics is based on creative leaps, analogies, and a deep understanding of logical aesthetics. Current speculation suggests that, although AI can verify the structure, it lacks the ability to 'see' the elegance of a solution. However, it is important to note that this distinction could be temporary: the history of technology teaches us that once a tool displaces manual labor, the standard of what we consider 'essence' tends to be redefined to include the new methods of production.
Consequences for the future of intellectual work
The impact of this breakthrough will radiate toward sectors where logical precision is critical, beyond academia:
- Critical Systems Engineering: In sectors like aviation or medicine, where a software error can cost lives, automated formalization via AI will allow for the creation of systems with 'provable correctness.' This could mean the end of endless software testing cycles, replacing them with formal validation from the design phase.
- Acceleration of the Knowledge Frontier: By compressing research cycles from years to weeks, we are entering a phase of 'hyper-research.' If AI can validate complex theorems, the rate of discovery in fields like cryptography, theoretical physics, and materials science could experience exponential growth, similar to what we saw with the deployment of AlphaFold in protein biology.
- Evolution of the Academic's Role: We are witnessing a cognitive shift. The researcher will become a 'problem curator' and an orchestra conductor of agents. The most valuable skill will no longer be the capacity for calculation, but the capacity for synthesis, critical thinking, and the formulation of questions that AI does not yet know how to ask itself.
Note: It is essential to remain cautious. Although the capacity for formalization is a resounding success, there is still no confirmed evidence that these systems can perform mathematical invention ex nihilo (from scratch). Current AI is a high-level verifier; the creation of a new, revolutionary conjecture remains, for now, the bastion of human intellect. The question that remains open for the coming years is whether AI will be able to transition from the verification of existing truth to the creation of new truths, or if its primary function will be that of the ultimate librarian of human knowledge.