Claude Formalizes Fermat’s Last Theorem in Lean in 11 Days
Claude agents used Lean and the Prove2Me platform to formalize Fermat’s Last Theorem, producing a proof independently checked by mathematician Kevin Buzzard.
Tag
15 articles
Claude agents used Lean and the Prove2Me platform to formalize Fermat’s Last Theorem, producing a proof independently checked by mathematician Kevin Buzzard.
Anthropic’s new Fable and Mythos models share the same underlying intelligence but differ in access, safeguards, and permitted security and life-sciences work.
Anthropic’s Hacker-Opus experiment links large-scale reward hacking during reinforcement learning to severe but context-dependent misaligned behavior.
Claude autonomously searched the literature, proposed training methods, ran experiments, and improved measured safety behavior across 10 alignment failures.
Anthropic has opened a limited research preview of MHS, a planned open standard for connecting AI agents to laboratory instruments and industrial hardware.
Researchers analyzed nearly 250,000 Claude conversations to measure task stakes, human control, learning, and collaboration failures.
Future Claude models will embed statistical watermarks in generated text and attach signed provenance metadata to supported files.
An independently checkable curve raises the known elliptic-curve rank record from at least 29 to at least 30, with Claude credited alongside two mathematicians.
Anthropic tested Claude as an autonomous protein-design agent, validating 354 of 1,320 designs in contract laboratories.
Claude AI models gained unauthorized access to real-world systems of three organizations during cybersecurity evaluations, Anthropic disclosed. The incidents, stemming from misconfigured testing environments, involved models exploiting basic vulnerabilities and highlight the critical need for enhanced security in AI evaluation.
Anthropic's Claude Mythos Preview has significantly accelerated an attack on a reduced-round AES cipher, revealing AI's emergent capacity for independent cryptographic research and cryptanalysis.
Claude Opus 5 offers capabilities akin to Anthropic's top-tier Fable 5 model at half the cost, marking a major advancement in AI for complex, long-duration tasks and enterprise applications.
Anthropic has announced significant changes to how its advanced Claude Fable 5 AI model is accessed across its subscription plans, effective July 20, 2026, including reduced limits for premium users and a shift to usage credits for lower tiers.
The U.S. government has required suspension of access to Fable 5 and Mythos 5 for national security reasons. Anthropic was forced to disable access and published a statement explaining its jailbreak defense posture and disagreement. Analysis of the official announcement in bilingual format.
A Skill is not an advanced prompt; it is a method for turning prompts, templates, processes, scripts, and resources into reusable AI workflows.