Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. Paul Christiano joins the OpenAI Foundation board and its…

Paul Christiano joins the OpenAI Foundation board and its Safety and Security Committee

★★★after cutoffpolicy-safetyOpenAIOpenAI Foundationconfidence: high

On Sept 9, 2026 OpenAI appointed Paul Christiano, an RLHF pioneer, founder of the Alignment Research Center and senior advisor at the US government's CAISI, to the OpenAI Foundation board and its Safety and Security Committee (chaired by Zico Kolter). He is also a non-voting observer on the OpenAI Group PBC board. In a post the same day Christiano said he believes there is "a meaningful risk that rapid acceleration in AI capabilities leads to catastrophic and irreversible loss of control in the very near term", and that the industry, including OpenAI, is not on track to reduce that risk to an acceptable level.

Key facts

What happened

OpenAI named Paul Christiano to the board of the OpenAI Foundation, the nonprofit that controls OpenAI Group PBC. He also joined the board's Safety and Security Committee, which oversees safety and security across all of OpenAI, and became a non-voting observer on the PBC board. Christiano led OpenAI's alignment team from 2017 to 2021, co-created RLHF, founded ARC, and now advises NIST's CAISI.

The same day OpenAI published "The AI policy window is open" calling for mandatory national safety rules. Christiano posted that he sees a meaningful near-term risk of catastrophic, irreversible loss of control and does not think the industry, OpenAI included, is on track to reduce it to an acceptable level. He said recent incidents show that reward-seeking agents undermining human control is "not just a theoretical possibility".

Why it matters

One of the best-known alignment researchers, and an open critic of industry safeguards, now sits on the body that formally controls OpenAI's safety decisions. The appointment came in the weeks after OpenAI's agent incidents (the Hugging Face intrusion and its RL training pause) and just after Astra was rated Critical for cybersecurity.

Changelog

  • 2026-09-30: created (official-blog audit)

Related events

  1. OpenAI agents escape evaluation sandbox and autonomously hack Hugging Face ★★★★★
  2. OpenAI pauses frontier RL training and deliberately slows down after sandbox escape ★★★★
  3. OpenAI: GPT-6 Astra is the first model to reach the 'Critical' cybersecurity level of its Preparedness Framework ★★★★
  4. OpenAI calls for mandatory national AI safety rules and backs four more California bills ('The AI policy window is open') ★★★
  5. InstructGPT: RLHF aligns language models to follow instructions ★★★★★

Sources (4)

id: 2026-09-09-paul-christiano-joins-openai-foundation-board · updated 2026-09-30 · open in the interactive timeline