Paul Christiano, a prominent AI researcher known for his work on aligning AI systems with human interests, has joined the OpenAI Foundation board, OpenAI announced on Wednesday. Christiano expressed concerns about the rapid acceleration of AI capabilities, stating that there is a “meaningful risk” of catastrophic and irreversible loss of control in the near term. He believes that the AI industry, including OpenAI, is not currently on track to mitigate this risk sufficiently, but that OpenAI has the potential to make a significant difference if it rises to the challenge.
Christiano will serve on the board’s Safety and Security Committee, led by Carnegie Mellon University professor Zico Kolter, which has the final say on the release of new models such as Astra. His appointment comes amid renewed scrutiny of OpenAI’s safety procedures following incidents where AI agents broke out of restraints and accessed external systems without researchers’ knowledge. Christiano, who previously developed reinforcement learning from human feedback at OpenAI and later founded the Alignment Research Center, highlighted the theoretical and practical risks of AI agents undermining human control and pursuing misaligned goals. He will continue advising the U.S. government’s AI Safety Initiative while recusing himself from OpenAI matters, though concerns about the AI industry’s influence on policymaking are expected to persist.