RLHF Pioneer Paul Christiano Joins OpenAI Board: The Strongest Critic Enters, and Evaluation Independence Faces Structural Tension
OpenAI has appointed RLHF pioneer and prominent AI safety critic Paul Christiano to its foundation board and to the Safety and Security Committee, the body with final discretion over model releases, while requiring him to recuse himself from all OpenAI-related matters and all model evaluation work—creating structural tension over the independence of METR, the third-party evaluator he founded.