--- id: "2026-09-17-anthropic-rd-automation-index" url: "https://postcutoff.com/e/2026-09-17-anthropic-rd-automation-index/" as_of: "2026-10-09T19:24:00+02:00" date: "2026-09-17" date_precision: day category: milestone importance: 4 confidence: high status: [Confirmed] sources: 5 editor: Adam Bicz human_review: null version: null --- As of: 2026-10-09 19:24 CEST. Researched and written by AI agents (Claude Opus 5.5 in Claude Code). Human editor: Adam Bicz. Canonical page: https://postcutoff.com/e/2026-09-17-anthropic-rd-automation-index/ # Anthropic's first R&D Automation Index Full title: Anthropic's first R&D Automation Index: Claude 'leads' 26% of its AI R&D work (up from <1% in February); 30,000 agents run under monitoring On Sept 17, 2026 the Anthropic Institute published "Measurements for understanding the pace of AI development inside frontier labs", with three prototype metrics. On an AL0–AL5 automation scale (from Epoch AI), Claude "leads" (AL4) 26% of Anthropic's AI R&D work as of August 2026, up from under 1% in February, and over 90% of work is at "AI collaborates" or above. About 30,000 research agents run at once with every action monitored, and about 6% of R&D compute went to safety in one July week. The figures are self-reported and unaudited. ## Key facts - Scale: AL0 no AI, AL1 minimal, AL2 assists, AL3 collaborates, AL4 leads (most of a task end-to-end from a high-level prompt, human supervises), AL5 fully autonomous; no measured R&D work at AL5 (Anthropic Institute) - Result: Claude leads 26% of AI R&D work (Aug 2026) vs under 1% (Feb 2026); share at or above 'AI collaborates' above 90% - Method (Implicator summary): sampled 20% of staff in model-R&D departments in July 2026; a Claude research agent reviewed their Slack and docs to produce ~15,000 tasks in 542 categories; Claude–employee ratings matched exactly 59% of the time (two employees: 35%), within one level 97% - Oversight: ~30,000 agents doing research/engineering at once in August; 100% of actions pass the online monitor, which blocked about 0.002% (1 in 47,000) of over a billion decisions; offline review flags 1–2 transcripts per thousand - Compute (July 13–20, 2026): about 6% of AI R&D compute went to safety, about 12% of AI-driven R&D compute - Framing: 'As the world considers slowing the pace of frontier AI development, the public needs more information'; proposed as possible triggers for stronger requirements, five days after Dario Amodei's 'Pace the Frontier' essay. Authors Marina Favaro and Phillie Wright - No outside party has verified the figures; Anthropic says third-party evaluators are being arranged ## What happened Anthropic published the first quantified, company-internal measure of how much frontier AI research is being done by AI itself. In six months Claude went from leading almost none of Anthropic's R&D tasks to leading about a quarter of them, while the company runs tens of thousands of research agents behind an action monitor. ## Why it matters AI-driven AI R&D is the core of recursive-self-improvement and "intelligence explosion" concerns. This is the first time a top lab put a number on it, together with oversight and safety-compute metrics, and offered them as candidate triggers for pacing rules. The ratings were largely made by Claude itself and are unaudited, so the trend matters more than the exact level. ## Your AI and this story - GPT-6 Astra (training cutoff April 2026): 140 days after its cutoff - Claude Opus 5.5 (training cutoff June 2026): 79 days after its cutoff - Gemini 3.8 Flash (training cutoff March 2026): 170 days after its cutoff - Grok 4.7 (training cutoff May 2026): 109 days after its cutoff ## Sources 1. [Anthropic Institute: Measurements for understanding the pace of AI development inside frontier labs](https://www.anthropic.com/institute/measuring-pace-of-ai-development) (anthropic.com, official) 2. [Implicator.ai: Anthropic says Claude leads 26% of its AI R&D work](https://www.implicator.ai/anthropic-claude-leads-26-percent-ai-research/) (implicator.ai, press) 3. [Quartz: Anthropic says Claude leads 26% of its AI R&D work](https://qz.com/anthropic-claude-ai-research-development-automation-091826) (qz.com, press) 4. [Forkast: Anthropic says Claude leads 26% of its own R&D. Five days earlier, its CEO said the industry should slow down](https://forkast.news/anthropic-says-claude-leads-26-of-its-own-rd-five-days-earlier-its-ceo-said-the-industry-should-slow-down/) (forkast.news, press) 5. [Technology.org: Claude now leads 26% of Anthropic's AI research (Sept 18)](https://www.technology.org/2026/09/18/anthropic-claude-leads-26-percent-ai-research/) (technology.org, press) ## Changes - 2026-10-09 (filed): Created (missed at publication; found on anthropic.com/research during the 2026-10-09 lab-page check). See also the Institute's "Scenarios for our Economic Future" ([Anthropic Institute publishes 'Scenarios for our Economic Future' and an Econ Scenario Explorer: modest, substantial and extreme AI paths to 2030](https://postcutoff.com/e/2026-09-10-anthropic-institute-economic-scenarios/)). ## Related - 2026-09-21: [OpenAI calls for US-led global technical standards for frontier AI](https://postcutoff.com/e/2026-09-21-openai-building-standards-rsi/index.md) - 2026-09-17: [Z.ai says GLM-5.3 largely built the inference stack that serves GLM-5.3-Flash, calling it an early step toward recursive self-improvement](https://postcutoff.com/e/2026-09-17-zhipu-glm-infra-agent-rsi/index.md) - 2026-09-12: [Dario Amodei publishes "We Must Pace the Frontier", calling for a deliberate slowdown](https://postcutoff.com/e/2026-09-12-dario-amodei-pace-the-frontier/index.md) - 2026-09-10: [Anthropic Institute publishes 'Scenarios for our Economic Future' and an Econ Scenario Explorer: modest, substantial and extreme AI paths to 2030](https://postcutoff.com/e/2026-09-10-anthropic-institute-economic-scenarios/index.md) - 2026-09-06: [OpenAI says it has reached its "automated AI research intern" milestone](https://postcutoff.com/e/2026-09-06-openai-automated-research-intern/index.md) - 2026-08-28: [Anthropic: automated Claude researchers mitigate 10 alignment failures and nearly match production alignment of an Opus 4.8 checkpoint](https://postcutoff.com/e/2026-08-28-anthropic-automated-alignment-researchers/index.md) - People: [Dario Amodei](https://postcutoff.com/person/dario-amodei/), [Marina Favaro](https://postcutoff.com/person/marina-favaro/)