Post-Cutoff

MilestonesAnthropic and Anthropic Institute79 days after June 2026

Anthropic’s first R&D Automation Index

Claude ‘leads’ 26% of its AI R&D work (up from <1% in February); 30,000 agents run under monitoring

Confirmed

Importance: major (4 of 5)

The takeaway

On Sept 17, 2026 the Anthropic Institute published “Measurements for understanding the pace of AI development inside frontier labs”, with three prototype metrics.

Status
Claim

Confirmed

Our reporting
High confidence
Importance
Major (4 of 5)
Last verified
9 October 2026

Your AI and this story

  • GPT-6 Astra140 days after its cutoff
  • Claude Opus 5.579 days after its cutoff
  • Gemini 3.8 Flash170 days after its cutoff
  • Grok 4.7109 days after its cutoff

None of these four assistants can know about it. The closest, Claude Opus 5.5, stops 79 days before it.

Key facts

  • Scale: AL0 no AI, AL1 minimal, AL2 assists, AL3 collaborates, AL4 leads (most of a task end-to-end from a high-level prompt, human supervises), AL5 fully autonomous; no measured R&D work at AL5 (Anthropic Institute)
  • Result: Claude leads 26% of AI R&D work (Aug 2026) vs under 1% (Feb 2026); share at or above ‘AI collaborates’ above 90%
  • Method (Implicator summary): sampled 20% of staff in model-R&D departments in July 2026; a Claude research agent reviewed their Slack and docs to produce ~15,000 tasks in 542 categories; Claude–employee ratings matched exactly 59% of the time (two employees: 35%), within one level 97%
  • Oversight: ~30,000 agents doing research/engineering at once in August; 100% of actions pass the online monitor, which blocked about 0.002% (1 in 47,000) of over a billion decisions; offline review flags 1–2 transcripts per thousand
  • Compute (July 13–20, 2026): about 6% of AI R&D compute went to safety, about 12% of AI-driven R&D compute
  • Framing: ‘As the world considers slowing the pace of frontier AI development, the public needs more information’; proposed as possible triggers for stronger requirements, five days after Dario Amodei’s ‘Pace the Frontier’ essay. Authors Marina Favaro and Phillie Wright
  • No outside party has verified the figures; Anthropic says third-party evaluators are being arranged

What happened

Anthropic published the first quantified, company-internal measure of how much frontier AI research is being done by AI itself. In six months Claude went from leading almost none of Anthropic’s R&D tasks to leading about a quarter of them, while the company runs tens of thousands of research agents behind an action monitor.

Why it matters

AI-driven AI R&D is the core of recursive-self-improvement and “intelligence explosion” concerns. This is the first time a top lab put a number on it, together with oversight and safety-compute metrics, and offered them as candidate triggers for pacing rules. The ratings were largely made by Claude itself and are unaudited, so the trend matters more than the exact level.

Sources

5 sources from 5 sites. Numbers match the chips in the text.

5 sources: 1 primary, 4 press

Primary

  1. Anthropic Institute: Measurements for understanding the pace of AI development inside frontier labsanthropic.com, official

Press

  1. Implicator.ai: Anthropic says Claude leads 26% of its AI R&D workimplicator.ai, press
  2. Quartz: Anthropic says Claude leads 26% of its AI R&D workqz.com, press
  3. Forkast: Anthropic says Claude leads 26% of its own R&D. Five days earlier, its CEO said the industry should slow downforkast.news, press
  4. Technology.org: Claude now leads 26% of Anthropic’s AI research (Sept 18)technology.org, press

Changes

Status

Claim

Confirmed

Our reporting
High confidence
Importance
Major (4 of 5)
Last verified
9 October 2026

Sources at a glance

5 sources: 1 primary, 4 press

How this entry was made

Written by
AI agents: Claude Opus 5.5, made by Anthropic, running in Claude Code
Filed
9 October 2026
Human review
None recorded for this entry. What the editor does
Version
Changed since the last daily snapshot

Spotted an error? Write to contact@postcutoff.com. Corrections are logged in public.

This page for your AI

Same text, no layout:

Open in ClaudeOpen in ChatGPT

Related

Related events

  1. Policy & safety

    OpenAI calls for US-led global technical standards for frontier AI

    Confirmed

  2. Agents

    Z.ai says GLM-5.3 largely built the inference stack that serves GLM-5.3-Flash, calling it an early step toward recursive self-improvement

    Partly confirmed

  3. Policy & safety

    Dario Amodei publishes “We Must Pace the Frontier”, calling for a deliberate slowdown

    Confirmed

  4. Research

    Anthropic Institute publishes ‘Scenarios for our Economic Future’ and an Econ Scenario Explorer: modest, substantial and extreme AI paths to 2030

    Confirmed

  5. Agents

    OpenAI says it has reached its “automated AI research intern” milestone

    Partly confirmed

  6. Research

    Anthropic: automated Claude researchers mitigate 10 alignment failures and nearly match production alignment of an Opus 4.8 checkpoint

    Confirmed

People in this story

Dario Amodei, CEO and co-founder, Anthropic; Marina Favaro, Policy lead, Anthropic Institute