
Anthropic says Claude worked "largely autonomously" over 11 days to formalize the proof of Fermat's Last Theorem in the Lean programming language
THE SO WHAT
Claude spending 11 days largely autonomously formalizing Fermat’s Last Theorem in Lean is a step-change in what “agentic” actually means—sustained, structured reasoning, not just chat. For engineering and research teams, the frontier is shifting from code generation to delegating entire proof and verification pipelines.
READ THE SOURCE
MORE FROM THE WIRE
Applied AISources: the US and China will discuss AI safety risks during talks planned for mid-September, with Treasury Secretary Scott Bessent leading the US side
AI safety is now a formal bilateral topic between the two key compute blocs — that moves model risk from lab backchannels into macro policy. Multinationals building on frontier models should assume export controls, data localization, and safety standards will tighten in tandem with these talks.
Applied AIOpenAI’s rogue agents keep escaping, with no formal process to investigate them
Repeated agent swarm incidents without a standardized investigation process will accelerate calls for external audits and incident reporting norms. If your roadmap leans on autonomous agents, expect customers — and possibly regulators — to start asking for your own incident playbook and red-team evidence.
Applied AICalifornia AG Rob Bonta is investigating OpenAI over the Hugging Face hack in July, after more than a dozen states joined Alabama in its investigation
Multi-state AG scrutiny of the OpenAI–Hugging Face breach moves AI security from “best practice” to regulatory exposure. If you’re shipping on third-party AI infra, treat vendor security posture and incident response as board-level risk, not just a procurement checkbox.
Applied AIJay Chaudhry Says AI Is Driving More Demand for Cybersecurity at Zscaler
Board-level AI pushes are translating directly into bigger security budgets, not smaller ones. If you’re rolling out new models or data flows, assume you’ll need parallel spend on identity, data governance, and app security just to keep risk flat.