Anthropic Says Claude Now Leads 26% Of Its Own AI R&D Work

Anthropic's inaugural R&D Automation Index puts Claude in the 'leads' seat for 26% of its own AI research, up from under 1% in February 2026, with 30,000 agents running in parallel on its internal platform.

Anthropic Says Claude Now Leads 26% Of Its Own AI R&D Work

Anthropic on September 18, 2026 published the first edition of its R&D Automation Index, an internal measurement designed to disclose how much of its own model-development work is now being driven by Claude. The headline number: Claude sits at the "leads" tier for 26% of Anthropic's AI R&D, up from under 1% as recently as February 2026.

An AL0-AL5 Ladder For AI-In-The-Loop

The index uses a six-step scale, from AL0 (no AI involvement) through AL5 (fully autonomous). Claude's headline share sits at AL4 – a level Anthropic defines as "the AI completes most tasks end-to-end from a high-level prompt, with human supervision." Anthropic is explicit that the top tier, AL5, is not yet reached: "Claude is not operating fully autonomously for any measured subset of AI R&D work."

Anthropic logo

30,000 Agents, 1 In 47,000 Blocked

Anthropic also disclosed the raw operational metrics behind the number. Its primary internal platform ran ~30,000 Claude agents simultaneously during August 2026, and its safety-monitoring system blocked roughly 1 in 47,000 agent decisions out of more than a billion total. Roughly 6% of AI R&D compute for the week of July 13–20, 2026 went to safety-focused work; that share rises to 12% when restricted to AI-driven R&D. Anthropic has committed to publishing the index regularly, with third-party verification.

Methodology In Brief

To score the index, Anthropic sampled 20% of staff across its model-R&D departments during July 2026. Claude agents then reviewed the sampled work and broke it into ~15,000 granular tasks arranged across 542 nodes, with human raters checking a subset. Model ratings and employee ratings agreed exactly 59% of the time. That agreement rate is Anthropic's answer to the obvious objection: an AI grading its own R&D deserves a control group.

Why It Matters

The disclosure is the most quantitatively specific look any frontier lab has given at how far AI-in-the-loop model development has actually progressed inside a top-tier lab. It also arrives at a moment when the same firms are asking regulators for wider latitude: it dovetails with the joint FINRA-style safety body Anthropic, OpenAI and Google DeepMind are pushing, and with Anthropic's own Bay Area wet-lab announcement earlier this week. Together those moves paint a picture of a lab arguing publicly that it is still safely in control while quietly pointing out that AI is doing more than a quarter of its own homework.

Reporting based on coverage from Implicator.ai, Enterprise DNA, Technology.org and Anthropic's own R&D Automation Index disclosure.

Category: AI & Technology

Tags: AI AI Foundation Models AI Agents agentic AI AI safety

Related Articles