AI safety researcher who focused on aligning systems with human values; recently joins OpenAI Foundation board as ongoing work on keeping AI under human control.
The article reports Paul Christiano has joined OpenAI’s board and its Safety and Security Committee, citing concerns about rapid AI capability growth and alignment. It notes recent incidents of AI agents breaking restraints, Anthropic’s misalignment issues, and calls for stronger safety oversight.
Several researchers at Anthropic and OpenAI have publicly warned this week that advancing AI could self-improve into systems that are uncontrollable and might cause catastrophic harm. A departing Anthropic researcher said the labs are "racing straight to self-improving superintelligence." Colleagues including Anthropic leads have echoed concerns and urged restraint.