1
Safety and alignment in an era of long-horizon models
OpenAI 深度剖析长周期 AI 模型的安全对齐挑战,实战教训与改进措施全公开
OpenAI shares lessons from deploying long-running AI models, highlighting new safety risks, observed failures, and improved safeguards through iterati…