Loading timeline…
20232020s
Organizations
DeepMindOpenAIAnthropic
Major Achievements
- •Co-developed the RLHF (Reinforcement Learning from Human Feedback) paper at OpenAI.
- •Co-led the Superalignment team to solve existential control problems for AGI within four years.
- •Resigned publicly from OpenAI over concerns that safety culture was being sidelined for shiny products.
Key Papers
- Deep Reinforcement Learning from Human Preferences
Paul F. Christiano et al. · NeurIPS 2017
- Training Language Models to Follow Instructions with Human Feedback
Long Ouyang · NeurIPS 2022