Third-party cyber evaluations involving OpenAI models
OpenAI explains recent third-party cybersecurity evaluation incidents and outlines new safeguards to strengthen AI model testing and evaluation.
OpenAI is investing in stronger safeguards and defensive capabilities as AI models become more powerful in cybersecurity. We explain how we assess risk, limit misuse, and work with the security community to strengthen cyber resilience.
OpenAI explains recent third-party cybersecurity evaluation incidents and outlines new safeguards to strengthen AI model testing and evaluation.
OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.
OpenAI outlines a five-part action plan for strengthening cybersecurity in the Intelligence Age, focused on democratizing AI-powered cyber defense and protecting critical systems.
Our goal is to facilitate the development of AI-powered cybersecurity capabilities for defenders through grants and other support.
The OpenAI Blog article titled "Path to Astra: critical capabilities and frontier safeguards" announces that Astra is the first OpenAI model to meet the Critical cybersecurity capability threshold under the Preparedness Framework. It highlights that Astra incorporates stronger safeguards for its release, ensuring higher security standards. The piece underscores the model’s compliance with stringent cybersecurity criteria.
arXiv:2607. 25379v1 Announce Type: new Abstract: Cyber-capable AI agents combine language models with tools, memory, and execution en- vironments to perform multi-step offensive-security tasks.
OpenAI’s mission is to build safe AI, and ensure AI’s benefits are as widely and evenly distributed as possible.
AI is reshaping cybersecurity for attackers and defenders alike. Learn how OpenAI is strengthening its defenses and what security teams can do now.
OpenAI has launched an initiative aimed at strengthening democratic oversight of artificial intelligence in national security. The program will provide government institutions with tools, training, and expertise to better manage AI’s role in security contexts.
OpenAI is enhancing monitoring, alignment, and security for frontier AI models. The company’s new safeguards are shaping how quickly these models are developed. This approach reflects a focus on responsible advancement of AI capabilities.
OpenAI expands its Trusted Access for Cyber program, introducing GPT-5. 4-Cyber to vetted defenders and strengthening safeguards as AI cybersecurity capabilities advance.