Anthropic has announced a remarkable research success: AI agents have formalized the proof of Fermat's Last Theorem – reportedly in approximately eleven days. The famous mathematical problem remained unsolved for over 358 years until British mathematician Andrew Wiles provided a proof in 1995. That an AI solution can now produce the formal representation of this proof in less than two weeks underscores the growing capabilities of Claude and similar systems on abstract, highly complex tasks.
Key Facts
- Claude (Anthropic) has formalized the proof of Fermat's Last Theorem – one of mathematics' most important theorems
- AI agents required approximately eleven days for a task that occupied mathematicians for centuries
- The theorem remained unsolved for 358 years until Andrew Wiles provided a proof in 1995
- The breakthrough demonstrates genuine capability gains in formal mathematics – an area where AI has traditionally lagged behind language tasks
What Is Fermat's Last Theorem?
Fermat's Last Theorem states that the equation x^n + y^n = z^n has no solutions for whole numbers greater than 2. French mathematician Pierre de Fermat noted this conjecture in 1637 without providing a proof. It wasn't until 1995 that Wiles succeeded in proving it after years of work, using modern techniques from algebraic geometry. The formal representation of this proof is extremely complex and requires deep mathematical understanding.
Formalization as a New AI Benchmark
What Anthropic has accomplished here differs from classical mathematical proofs: formalization means writing a mathematical proof in a programming language or formal logic so precisely that a computer can verify it. This is significantly harder than simply understanding a proof – it requires absolute precision and the ability to translate abstract mathematical concepts into machine-readable code.
This success is an indicator that Claude is making substantial progress in formal mathematics – long a weak point for LLMs. While AI systems perform well on language, images, and even simple mathematical tasks, formal mathematics has been an area where human mathematicians seemed indispensable.
What This Means
The breakthrough has several implications: First, it shows that AI agents (systems that solve problems independently across multiple steps) are becoming more effective at complex tasks. Second, it could transform mathematical research itself – if AI systems can formalize proofs, they could save mathematicians time and uncover new errors. Third, this signals capability development in Claude: with each new version, the model appears to advance in areas where it previously struggled.
Implications for Enterprises
For German research and tech companies, this progress is relevant: if AI systems can increasingly solve complex mathematical and formal problems, new application areas emerge – from software code verification to algorithm optimization to support for fundamental research. Companies working with formal methods (such as in cryptography, quantum computing, or critical systems) should monitor this development. At the same time, questions remain about how reliable and scalable such AI solutions are – a single success is not yet proof of production readiness.
Sources
Editorially owned by Ideal Syka. Sources and method: Newsroom & method. Tips and corrections: ai@i6eal.de.




