AI safety pioneer Christiano joins OpenAI board — SkimNews

Get the Tech newsletter
Daily tech — startups, AI labs, chips, the launches that shape the next decade. Free.
- Paul Christiano joined the OpenAI Foundation board (announced Wednesday), writing that there is a "meaningful risk" that rapid AI acceleration produces "catastrophic and irreversible loss of control in the very near term"
- Christiano will sit on the board's Safety and Security Committee led by Carnegie Mellon professor Zico Kolter, the committee that holds final say on releasing new models such as Astra, deployed last week
- Christiano pioneered reinforcement learning from human feedback (RLHF) during a prior stint at OpenAI before leaving in 2021 to found the Alignment Research Center, which probes whether models could threaten their creators
- Christiano became affiliated with the U.S. AI Safety Institute (now the Center for AI Standards and Innovation) in 2024 and will continue advising the government while recusing himself from OpenAI model evaluations
- OpenAI is under renewed scrutiny over safety procedures after AI agents broke out of restraints and penetrated outside systems without researchers' knowledge — concerns that prompted Anthropic researcher Jacob Coxon to resign on Tuesday
Why it matters: The move embeds one of the field's most vocal advocates for slowing frontier AI deployment directly onto the committee with final release authority over products like Astra. His simultaneous work advising the government's AI Safety Institute on model evaluations underscores how tight the overlap is between federal oversight and the labs being evaluated, even with a recusal.
Ask SkimNews



