An update on our safety & security practices
An update on our safety & security practices
Related stories
Helping people when they need it most
How we think about safety for users experiencing mental or emotional distress, the limits of today’s systems, and the work underway to refine them.
Operator System Card
Drawing from OpenAI’s established safety frameworks, this document highlights our multi-layered approach, including model and product mitigations we’ve implemented to protect against prompt engineering and jailbreaks, protect privacy and security, as well as details our external red teaming efforts, safety evaluations, and ongoing work to further refine these safeguards.
Our updated Preparedness Framework
Sharing our updated framework for measuring and protecting against severe harm from frontier AI capabilities.
OpenAI Board Forms Safety and Security Committee
Our commitment to community safety
Learn how OpenAI protects community safety in ChatGPT through model safeguards, misuse detection, policy enforcement, and collaboration with safety experts.
Frontier Model Forum
We’re forming a new industry body to promote the safe and responsible development of frontier AI systems: advancing AI safety research, identifying best practices and standards, and facilitating information sharing among policymakers and industry.
OpenAI’s commitment to child safety: adopting safety by design principles
Why responsible AI development needs cooperation on safety
We’ve written a policy research paper identifying four strategies that can be used today to improve the likelihood of long-term industry cooperation on safety norms in AI: communicating risks and benefits, technical collaboration, increased transparency, and incentivizing standards. Our analysis shows that industry cooperation on safety will be instrumental in ensuring that AI systems are safe and beneficial, but competitive pressures could lead to a collective action problem, potentially causing AI companies to under-invest in safety.
SafetyKit scales risk agents with OpenAI’s most capable models
Discover how SafetyKit leverages OpenAI GPT-5 to enhance content moderation, enforce compliance, and outpace legacy safety systems with greater accuracy .
Our approach to AI safety
Ensuring that AI systems are built, deployed, and used safely is critical to our mission.
Towards safety cases for frontier AI training
The article presents early guidelines for safety cases in frontier AI training, outlining technical safeguards, operational practices, and methods for investigating misalignment incidents. It emphasizes the importance of structured safety documentation to guide the development and deployment of advanced AI systems. The guidelines aim to provide a framework for identifying and mitigating risks associated with frontier AI training.