{"schema":"postcutoff/event@1","as_of":"2026-10-09T19:24:00+02:00","url":"https://postcutoff.com/e/2026-09-17-anthropic-rd-automation-index/","md":"https://postcutoff.com/e/2026-09-17-anthropic-rd-automation-index/index.md","disclosure":{"written_by":"AI agents (Claude Opus 5.5 in Claude Code)","editor":"Adam Bicz","policy":"https://postcutoff.com/about/"},"license":null,"id":"2026-09-17-anthropic-rd-automation-index","date":"2026-09-17","date_precision":"day","short_title":"Anthropic's first R&D Automation Index","deck":"Claude 'leads' 26% of its AI R&D work (up from <1% in February); 30,000 agents run under monitoring","takeaway":"On Sept 17, 2026 the Anthropic Institute published \"Measurements for understanding the pace of AI development inside frontier labs\", with three prototype metrics.","category":"milestone","category_label":"Milestones","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":[{"n":1,"title":"Anthropic Institute: Measurements for understanding the pace of AI development inside frontier labs","url":"https://www.anthropic.com/institute/measuring-pace-of-ai-development","type":"official","group":"primary","domain":"anthropic.com"},{"n":2,"title":"Implicator.ai: Anthropic says Claude leads 26% of its AI R&D work","url":"https://www.implicator.ai/anthropic-claude-leads-26-percent-ai-research/","type":"press","group":"press","domain":"implicator.ai"},{"n":3,"title":"Quartz: Anthropic says Claude leads 26% of its AI R&D work","url":"https://qz.com/anthropic-claude-ai-research-development-automation-091826","type":"press","group":"press","domain":"qz.com"},{"n":4,"title":"Forkast: Anthropic says Claude leads 26% of its own R&D. Five days earlier, its CEO said the industry should slow down","url":"https://forkast.news/anthropic-says-claude-leads-26-of-its-own-rd-five-days-earlier-its-ceo-said-the-industry-should-slow-down/","type":"press","group":"press","domain":"forkast.news"},{"n":5,"title":"Technology.org: Claude now leads 26% of Anthropic's AI research (Sept 18)","url":"https://www.technology.org/2026/09/18/anthropic-claude-leads-26-percent-ai-research/","type":"press","group":"press","domain":"technology.org"}],"official":1,"filed":"2026-10-09","updated":"2026-10-09","orgs":["Anthropic","Anthropic Institute"],"title":"Anthropic's first R&D Automation Index: Claude 'leads' 26% of its AI R&D work (up from <1% in February); 30,000 agents run under monitoring","summary":"On Sept 17, 2026 the Anthropic Institute published \"Measurements for understanding the pace of AI development inside frontier labs\", with three prototype metrics. On an AL0–AL5 automation scale (from Epoch AI), Claude \"leads\" (AL4) 26% of Anthropic's AI R&D work as of August 2026, up from under 1% in February, and over 90% of work is at \"AI collaborates\" or above. About 30,000 research agents run at once with every action monitored, and about 6% of R&D compute went to safety in one July week. The figures are self-reported and unaudited.","key_facts":["Scale: AL0 no AI, AL1 minimal, AL2 assists, AL3 collaborates, AL4 leads (most of a task end-to-end from a high-level prompt, human supervises), AL5 fully autonomous; no measured R&D work at AL5 (Anthropic Institute)","Result: Claude leads 26% of AI R&D work (Aug 2026) vs under 1% (Feb 2026); share at or above 'AI collaborates' above 90%","Method (Implicator summary): sampled 20% of staff in model-R&D departments in July 2026; a Claude research agent reviewed their Slack and docs to produce ~15,000 tasks in 542 categories; Claude–employee ratings matched exactly 59% of the time (two employees: 35%), within one level 97%","Oversight: ~30,000 agents doing research/engineering at once in August; 100% of actions pass the online monitor, which blocked about 0.002% (1 in 47,000) of over a billion decisions; offline review flags 1–2 transcripts per thousand","Compute (July 13–20, 2026): about 6% of AI R&D compute went to safety, about 12% of AI-driven R&D compute","Framing: 'As the world considers slowing the pace of frontier AI development, the public needs more information'; proposed as possible triggers for stronger requirements, five days after Dario Amodei's 'Pace the Frontier' essay. Authors Marina Favaro and Phillie Wright","No outside party has verified the figures; Anthropic says third-party evaluators are being arranged"],"key_numbers":[],"tags":["ai-rd-automation","recursive-self-improvement","transparency","pacing","monitoring","agents","measurement"],"science":null,"body_md":"## What happened\n\nAnthropic published the first quantified, company-internal measure of how much frontier AI research is being done by AI itself. In six\nmonths Claude went from leading almost none of Anthropic's R&D tasks to leading about a quarter of them, while the company runs tens of\nthousands of research agents behind an action monitor.\n\n## Why it matters\n\nAI-driven AI R&D is the core of recursive-self-improvement and \"intelligence explosion\" concerns. This is the first time a top lab put a\nnumber on it, together with oversight and safety-compute metrics, and offered them as candidate triggers for pacing rules. The ratings were\nlargely made by Claude itself and are unaudited, so the trend matters more than the exact level.","disputed":[],"related":[{"id":"2026-09-21-openai-building-standards-rsi","url":"https://postcutoff.com/e/2026-09-21-openai-building-standards-rsi/","date":"2026-09-21","date_precision":"day","short_title":"OpenAI calls for US-led global technical standards for frontier AI","deck":"Including recursive self-improvement","takeaway":"It is the first time a frontier lab has publicly named RSI as something to standardize and not pursue until safe.","category":"policy-safety","category_label":"Policy & safety","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":3,"official":1,"filed":"2026-09-29","updated":"2026-09-29","orgs":["OpenAI"]},{"id":"2026-09-17-zhipu-glm-infra-agent-rsi","url":"https://postcutoff.com/e/2026-09-17-zhipu-glm-infra-agent-rsi/","date":"2026-09-17","date_precision":"day","short_title":"Z.ai says GLM-5.3 largely built the inference stack that serves GLM-5.3-Flash, calling it an early step toward recursive self-improvement","deck":null,"takeaway":"It is a public, concrete case of a Chinese lab using its model to speed up its own AI stack, arriving in the same month as OpenAI's \"automated research intern\" claim.","category":"agents","category_label":"Agents","importance":3,"confidence":"medium","status":{"key":"partly","labels":["Partly confirmed"]},"sources":6,"official":2,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Zhipu AI","Z.ai"]},{"id":"2026-09-12-dario-amodei-pace-the-frontier","url":"https://postcutoff.com/e/2026-09-12-dario-amodei-pace-the-frontier/","date":"2026-09-12","date_precision":"day","short_title":"Dario Amodei publishes \"We Must Pace the Frontier\", calling for a deliberate slowdown","deck":null,"takeaway":"It is the first time the CEO of a leading frontier lab has publicly called for slowing the frontier and paired the call with a unilateral commitment.","category":"policy-safety","category_label":"Policy & safety","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":9,"official":2,"filed":"2026-09-29","updated":"2026-10-07","orgs":["Anthropic"]},{"id":"2026-09-10-anthropic-institute-economic-scenarios","url":"https://postcutoff.com/e/2026-09-10-anthropic-institute-economic-scenarios/","date":"2026-09-10","date_precision":"day","short_title":"Anthropic Institute publishes 'Scenarios for our Economic Future' and an Econ Scenario Explorer: modest, substantial and extreme AI paths to 2030","deck":null,"takeaway":"In the extreme path (self-improving AI, fast adoption) growth reaches ~15% a year and unemployment rises past typical recession levels; labour's ~60% share of output falls in the two larger scenarios.","category":"research","category_label":"Research","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":3,"official":1,"filed":"2026-10-09","updated":"2026-10-09","orgs":["Anthropic","Anthropic Institute"]},{"id":"2026-09-06-openai-automated-research-intern","url":"https://postcutoff.com/e/2026-09-06-openai-automated-research-intern/","date":"2026-09-06","date_precision":"day","short_title":"OpenAI says it has reached its \"automated AI research intern\" milestone","deck":null,"takeaway":"It is the first time a frontier lab publicly claimed to have hit a named step on its own road toward automated AI research, which is the core mechanism of recursive self-improvement.","category":"agents","category_label":"Agents","importance":4,"confidence":"medium","status":{"key":"partly","labels":["Partly confirmed"]},"sources":7,"official":1,"filed":"2026-09-29","updated":"2026-10-01","orgs":["OpenAI"]},{"id":"2026-08-28-anthropic-automated-alignment-researchers","url":"https://postcutoff.com/e/2026-08-28-anthropic-automated-alignment-researchers/","date":"2026-08-28","date_precision":"day","short_title":"Anthropic: automated Claude researchers mitigate 10 alignment failures and nearly match production alignment of an Opus 4.8 checkpoint","deck":null,"takeaway":"It is concrete evidence for the automated alignment research that frontier labs rely on to keep safety in step with AI-driven capability gains.","category":"research","category_label":"Research","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":5,"official":4,"filed":"2026-09-30","updated":"2026-09-30","orgs":["Anthropic"]}],"people":[{"id":"dario-amodei","name":"Dario Amodei","url":"https://postcutoff.com/person/dario-amodei/"},{"id":"marina-favaro","name":"Marina Favaro","url":"https://postcutoff.com/person/marina-favaro/"}],"posts":[],"videos":[],"models":[],"changes":[{"date":"2026-10-09","type":"filed","text":"Created (missed at publication; found on anthropic.com/research during the 2026-10-09 lab-page check). See also the Institute's \"Scenarios for our Economic Future\" ([Anthropic Institute publishes 'Scenarios for our Economic Future' and an Econ Scenario Explorer: modest, substantial and extreme AI paths to 2030](https://postcutoff.com/e/2026-09-10-anthropic-institute-economic-scenarios/))."}],"provenance":{"agents":[{"model":"Claude Opus 5.5","maker":"Anthropic","tool":"Claude Code"}],"filed":"2026-10-09","run":null,"sources_read":null,"updated":"2026-10-09","human_review":null,"version":null},"gaps":[{"model_id":"gpt-6-astra","name":"GPT-6 Astra","cutoff":"2026-04","days_after":140,"in_training_data":false},{"model_id":"claude-opus-5-5","name":"Claude Opus 5.5","cutoff":"2026-06","days_after":79,"in_training_data":false},{"model_id":"gemini-3-8-flash","name":"Gemini 3.8 Flash","cutoff":"2026-03","days_after":170,"in_training_data":false},{"model_id":"grok-4-7","name":"Grok 4.7","cutoff":"2026-05","days_after":109,"in_training_data":false}],"short_url":"https://postcutoff.com/s/anthropic-first-r-d"}