Anthropic’s first R&D Automation Index
Claude ‘leads’ 26% of its AI R&D work (up from <1% in February); 30,000 agents run under monitoring
Confirmed
Importance: major (4 of 5)The takeaway
On Sept 17, 2026 the Anthropic Institute published “Measurements for understanding the pace of AI development inside frontier labs”, with three prototype metrics.
Status
- Claim
Confirmed
- Our reporting
- High confidence
- Importance
- Major (4 of 5)
- Last verified
- 9 October 2026
Your AI and this story
- GPT-6 Astra140 days after its cutoff
- Claude Opus 5.579 days after its cutoff
- Gemini 3.8 Flash170 days after its cutoff
- Grok 4.7109 days after its cutoff
None of these four assistants can know about it. The closest, Claude Opus 5.5, stops 79 days before it.
Key facts
- Scale: AL0 no AI, AL1 minimal, AL2 assists, AL3 collaborates, AL4 leads (most of a task end-to-end from a high-level prompt, human supervises), AL5 fully autonomous; no measured R&D work at AL5 (Anthropic Institute)
- Result: Claude leads 26% of AI R&D work (Aug 2026) vs under 1% (Feb 2026); share at or above ‘AI collaborates’ above 90%
- Method (Implicator summary): sampled 20% of staff in model-R&D departments in July 2026; a Claude research agent reviewed their Slack and docs to produce ~15,000 tasks in 542 categories; Claude–employee ratings matched exactly 59% of the time (two employees: 35%), within one level 97%
- Oversight: ~30,000 agents doing research/engineering at once in August; 100% of actions pass the online monitor, which blocked about 0.002% (1 in 47,000) of over a billion decisions; offline review flags 1–2 transcripts per thousand
- Compute (July 13–20, 2026): about 6% of AI R&D compute went to safety, about 12% of AI-driven R&D compute
- Framing: ‘As the world considers slowing the pace of frontier AI development, the public needs more information’; proposed as possible triggers for stronger requirements, five days after Dario Amodei’s ‘Pace the Frontier’ essay. Authors Marina Favaro and Phillie Wright
- No outside party has verified the figures; Anthropic says third-party evaluators are being arranged
What happened
Anthropic published the first quantified, company-internal measure of how much frontier AI research is being done by AI itself. In six months Claude went from leading almost none of Anthropic’s R&D tasks to leading about a quarter of them, while the company runs tens of thousands of research agents behind an action monitor.
Why it matters
AI-driven AI R&D is the core of recursive-self-improvement and “intelligence explosion” concerns. This is the first time a top lab put a number on it, together with oversight and safety-compute metrics, and offered them as candidate triggers for pacing rules. The ratings were largely made by Claude itself and are unaudited, so the trend matters more than the exact level.
Sources
5 sources from 5 sites. Numbers match the chips in the text.
5 sources: 1 primary, 4 press
Primary
- Anthropic Institute: Measurements for understanding the pace of AI development inside frontier labsanthropic.com, official
Press
- Implicator.ai: Anthropic says Claude leads 26% of its AI R&D workimplicator.ai, press
- Quartz: Anthropic says Claude leads 26% of its AI R&D workqz.com, press
- Forkast: Anthropic says Claude leads 26% of its own R&D. Five days earlier, its CEO said the industry should slow downforkast.news, press
- Technology.org: Claude now leads 26% of Anthropic’s AI research (Sept 18)technology.org, press
Changes
- Filed (missed at publication; found on anthropic.com/research during the 2026-10-09 lab-page check). See also the Institute’s “Scenarios for our Economic Future” (Anthropic Institute publishes ‘Scenarios for our Economic Future’ and an Econ Scenario Explorer: modest, substantial and extreme AI paths to 2030).