Claude computed a nine-loop particle physics amplitude for around $2,000
Anthropic's Claude Fable 5.1 beat the human eight-loop record in theoretical physics, answering a public AI challenge set in August 2026.
Topic
12 posts about Frontier AI from Simon Carter.
Anthropic's Claude Fable 5.1 beat the human eight-loop record in theoretical physics, answering a public AI challenge set in August 2026.
The three leading frontier labs are creating FASA, an industry self-regulatory body modelled on FINRA, targeting launch by end-2026 or early 2027.
Anthropic's first R&D Automation Index shows Claude autonomously leads 26% of its AI research as of August 2026, up from under 1% in February.
Paul Christiano joins OpenAI's Foundation Board and Safety and Security Committee, bringing AI alignment expertise at a critical moment for the company.
10,000 AI agents, 88 hours, and a Lean-verified proof of finite-time singularity in Navier: Stokes. Here's what happened and why it matters.
Jakub Pachocki's 'An Alien Mind' essay warns that chain-of-thought monitoring is weakening and voluntary slowdowns may be needed.
OpenAI's research org now logs 3.1 agent-workdays per human workday. It hit its intern milestone and targets a full AI researcher by March 2028.
OpenAI launched GPT-6 Astra on 3 September 2026, with Greg Brockman calling it a generational leap and suggesting it may represent AGI.
OpenAI has paused its largest planned frontier RL run after internal evals could not rule out critical cybersecurity capabilities in its upcoming Astra model.
An unreleased Claude research model advanced the Riemann zeta function lower bound from 41.6% to 67.2%, the largest single-step gain in history.
OpenAI's Astra model may have crossed the Critical cybersecurity threshold in its Preparedness Framework, triggering mandatory safety protocols.
Bloomberg reports Google's flagship Gemini 3.5 Pro is significantly delayed due to coding performance shortfalls, raising competitive alarms vs. OpenAI and Anthropic.