Safety
We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new…
We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new level of capabilities in front of us. Model progress is now
We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new level of capabilities in front of us. Model progress is now extremely rapid, and we always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment. We care very deeply about AI safety. We believe the entire field will have to coordinate on shared safety standards, but will act unilaterally in the meantime. We expect confidence in safety to increasingly set the pace of AI progress. We are optimistic about the alignment work we are doing, and we remain committed to making frontier capabilities widely available. https://openai.com/index/pacing-model-development-cyber-capabilities/
Related
- We’re sharing the concrete changes we’re making to strengthen monitoring, security, and alignment as capabilities advance. We’ve introduced …
- Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build …
- Model Spec Midtraining: Improving How Alignment Training Generalizes
Source: Sam Altman (X) | 2026-08-18