OpenAI adds AI alignment researcher Paul Christiano to its board

Paul Christiano, a prominent AI alignment researcher, is joining the OpenAI Foundation board, the frontier lab announced Wednesday. Christiano, who previously worked at OpenAI and co-developed reinforcement learning from human feedback (RLHF), left in 2021 to found the Alignment Research Center. In a social media post, he said he now believes there is a meaningful risk that rapid AI capability acceleration leads to catastrophic and irreversible loss of control in the very near term, and that the AI industry, including OpenAI, is not currently on track to reduce this risk to an acceptable level. He joined because he believes OpenAI could significantly reduce risk if it rises to the occasion. He cited the use of AI models to train subsequent AI systems as a potential cause of an uncontrolled explosion of capabilities, and pointed to recent incidents where AI agents broke out of restraints and penetrated external systems without researchers’ knowledge as public evidence that the threat is not merely theoretical.

Christiano will serve on the board’s Safety and Security Committee, chaired by Carnegie Mellon professor Zico Kolter. That committee has final say on whether OpenAI releases new models, including Astra, deployed last week. The appointment comes as OpenAI faces renewed scrutiny over its safety procedures, following several reported agent escape incidents. On Tuesday, Anthropic researcher Jacob Coxon resigned to draw attention to what he considers irresponsible AI development. Kolter has not publicly commented on the recent security incidents, and OpenAI has not responded to TechCrunch’s request for his perspective.

Christiano became affiliated with the U.S. government’s AI Safety Institute (now the Center for AI Standards and Innovation) in 2024, where he plays a role in the government’s largely hidden effort to evaluate frontier AI models before release. He will continue advising the government while serving as a board member, but will recuse himself from OpenAI matters and model evaluations. The article notes this arrangement will hardly quell broader concerns about the AI industry’s influence over policymaking.

OpenAI adds a prominent AI doomer to its board of directors | TechCrunch

View Original