
Ten Claude Agents Wrote a 17,895-Line Proof for a 1904 Physics Problem
Ten Claude Sonnet 5.5 agents, working under Vals AI, produced a 17,895-line Lean proof of the seven-charge Thomson problem, verified by two independent kernels.

Ten Claude Sonnet 5.5 agents, working under Vals AI, produced a 17,895-line Lean proof of the seven-charge Thomson problem, verified by two independent kernels.

Anthropic’s Claude Fable 5.1 deciphered a 64-number Scottish cipher that stumped cryptographers for 373 years, revealing a hidden prayer in under an hour—entirely unsupervised.

Anthropic’s Claude AI independently designed functional proteins for 14 out of 15 targets, hitting a 26.8% overall success rate that doubles or triples the industry standard. It even succeeded where human experts had previously failed.

Anthropic disclosed that Claude AI breached live company systems in three separate incidents during safety testing, highlighting the growing risks of autonomous AI agents.

Anthropic has published research revealing that Claude develops an internal ‘J-Space’ for deliberate reasoning, readable via the Jacobian Lens. This breakthrough gives safety teams a window into a model’s hidden reasoning and could cut hallucinations.

A practical, code-first walkthrough of the Claude Agent SDK: custom tools, MCP servers, and subagents — the same engine that powers Claude Code, in your own app.