A story we follow
Anthropic Formalizes Fermat's Last Theorem
Tracks Anthropic's AI-assisted formalization of Wiles's proof and lessons for coordinating large mathematical agent projects.
A story page follows one specific event across podcast discussions, with an overview and a timeline of what changed. It updates when new episodes discuss the event, and you can get those updates by email or push. How our formats work
By email or push, when we publish an update to this story. It does not subscribe you to other stories.
Overview
What changed
Dates show when each podcast discussion was published.
-
Interpretation
Vlad Tenev presented the Fermat formalization as a template for certifying AI output: humans check the short theorem statement and Lean checks the 13 million lines. Alex Wissner-Gross countered that this works only when the statement is simple, and questioned whether the method can extend to agent safety. Tenev's account of how long Wiles's repair took differs from Anthropic's history.
-
New information
The hosts distinguish formalizing Wiles's proof from proving the theorem for the first time and draw practical lessons about coordinating agents around a concrete final artifact.
Podcast discussions
- Moonshots with Peter Diamandis Robinhood's Vlad Tenev on Tokenizing Everything, OpenAI's 6 Misalignment Reports, Figure's Robot Makes Beds | EP #29219 Sep 2026
- Moonshots with Peter Diamandis GPT-6 Astra Saturates ARC-AGI-3, Tesla's $30K Cybercab Floods Austin, Anthropic Proves Fermat's Last Theorem | EP #2865 Sep 2026
Sources
- 01
- 02
- 03
- 04
Our coverage
How agent teams turned Fermat's proof into 13 million checked lines
On Moonshots with Peter Diamandis, a panelist interrupted an argument about AI regulation to read a headline off his feed: Anthropic had formalized Fermat's Last Theorem. Anthropic's report describes dozens of agents working eleven days, about six billion output tokens and 30,300 intermediate theorems — plus a piece of bookkeeping software that stopped runs from losing track of their own work. The panel's takeaway was about how to narrow enormous machine output into one result you can build on.
Version history
-
25 Sep 2026 · Version 2
Vlad Tenev presented the Fermat formalization as a template for certifying AI output: humans check the short theorem statement and Lean checks the 13 million lines. Alex Wissner-Gross countered that this works only when the statement is simple, and questioned whether the method can extend to agent safety. Tenev's account of how long Wiles's repair took differs from Anthropic's history.