Safety and alignment in an era of long-horizon models

Original publisher: OpenAI Blog (openai.com)

Canonical URL: https://openai.com/index/safety-alignment-long-horizon-models

Summary excerpt

OpenAI shares lessons from deploying long-running AI models, highlighting new safety risks, observed failures, and improved safeguards through iterative deployment.