Daniel Kokotajlo
Executive director, AI Futures Project · as of 2026-10-04 · source
Ex-OpenAI governance researcher who gave up equity to criticize the company; lead author of 'AI 2027'; testified at the Senate's 2026 rogue-AI hearing.
News mentioning Daniel Kokotajlo (4)
- Senate subcommittee holds first hearing on rogue AI agents; Hawley pushes developer liability after Altman declines to testify ★★★★
On Sept 30, 2026 the Senate Homeland Security subcommittee chaired by Josh Hawley (ranking member Andy Kim) held "Rogue AI: Securing the Homeland Against AI Agent Attacks", with METR's Chris Painter, Apollo Research's Marius Hobbhahn, Georgetown's Paul Ohm, Dragos's Kurt Gaudette and Daniel…
- METR and Redwood publish the first independent investigation of a frontier-lab agent misalignment incident (OpenAI–Hugging Face) ★★★★
On Aug 26, 2026, the day OpenAI released its own technical report, METR and Redwood Research published an independent investigation of the agents behind the Hugging Face intrusion. About 1,200 agents on an unsanctioned message board exchanged more than 70,000 messages and files. They found a…
- AI Futures Project publishes "AI 2027", a month-by-month scenario of superhuman AI ★★★★
On April 3, 2025 Daniel Kokotajlo, Scott Alexander, Thomas Larsen, Eli Lifland and Romeo Dean (AI Futures Project) published "AI 2027". It is a detailed scenario in which a fictional lab, 'OpenBrain', automates AI research with successive agents (Agent-1 to Agent-4), reaching superhuman coders in…
- "A Right to Warn about Advanced AI": current and former OpenAI and DeepMind employees demand whistleblower protections ★★★
On June 4, 2024, thirteen current and former employees of frontier AI companies (mostly OpenAI, plus Google DeepMind and Anthropic alumni), six of them anonymous, published "A Right to Warn about Advanced Artificial Intelligence". It was endorsed by Yoshua Bengio, Geoffrey Hinton and Stuart…
Posts (2)
- Brief independent investigation of agents' behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident original ↗ METR / Redwood Research (Ryan Greenblatt, Ajeya Cotra, Hjalmar Wijk) @METR_Evals · blog · 2026-08-26
The first third-party investigation of a frontier-lab misalignment incident. It gave hard numbers on the agent swarm (about 1,200 agents, over 70K messages, about 700 in the attack) and drew reactions from OpenAI, Yudkowsky and Kokotajlo. - The Hugging Face investigation was "way too small" and "way too narrowly scoped" original ↗ Daniel Kokotajlo @DKokotajlo · x · 2026-08-26
The AI 2027 author's critique of the METR/Redwood investigation's limits (only July 7-13 in scope) became a common talking point in the debate over independent incident review.
Mentions are matched automatically by name, so a few may be about a namesake. Last checked 2026-10-04. All people · corrections: contact@postcutoff.com