Our updated Preparedness Framework
Sharing our updated framework for measuring and protecting against severe harm from frontier AI capabilities.
We’re strengthening the Frontier Safety Framework (FSF) to help identify and mitigate severe risks from advanced AI models.
Sharing our updated framework for measuring and protecting against severe harm from frontier AI capabilities.
We’re forming a new industry body to promote the safe and responsible development of frontier AI systems: advancing AI safety research, identifying best practices and standards, and facilitating information sharing among policymakers and industry.
To support the safety of highly-capable AI systems, we are developing our approach to catastrophic risk preparedness, including building a Preparedness team and launching a challenge.
The article presents early guidelines for safety cases in frontier AI training, outlining technical safeguards, operational practices, and methods for investigating misalignment incidents. It emphasizes the importance of structured safety documentation to guide the development and deployment of advanced AI systems. The guidelines aim to provide a framework for identifying and mitigating risks associated with frontier AI training.
Ensuring that AI systems are built, deployed, and used safely is critical to our mission.
OpenAI outlines a blueprint for U. S.
Explore OpenAI’s Frontier Governance Framework and how our AI safety, security, and risk practices align with emerging EU and California regulations.
OpenAI outlines priorities and principles for rigorous, secure, and independent third-party AI safety assessments of frontier models and safeguards.
arXiv:2607. 16112v1 Announce Type: new Abstract: Frontier AI companies have published capability thresholds that differ substantially, making it difficult for third parties to verify whether a threshold has been crossed or to compare requirements across companies.
OpenAI shares lessons from deploying long-running AI models, highlighting new safety risks, observed failures, and improved safeguards through iterative deployment.