OpenAI Blog

Frontier risk and preparedness

To support the safety of highly-capable AI systems, we are developing our approach to catastrophic risk preparedness, including building a Preparedness team and launching a challenge.

arXiv AI
Jul 17

Unsafe at any AUC: Unlearned Lessons from Sociotechnical Disasters for Responsible AI

arXiv:2607. 14353v1 Announce Type: cross Abstract: As automated decision-making and data-driven technologies pervade society and are used to manage consequential outcomes, understanding the technology's capabilities, limitations, and attendant risks in context requires analysis of full sociotechnical systems.

By Joshua A. Kroll, Andrew Smart, R. Stuart Geiger, Abigail Z. Jacobs
OpenAI Blog
5d ago

Towards safety cases for frontier AI training

The article presents early guidelines for safety cases in frontier AI training, outlining technical safeguards, operational practices, and methods for investigating misalignment incidents. It emphasizes the importance of structured safety documentation to guide the development and deployment of advanced AI systems. The guidelines aim to provide a framework for identifying and mitigating risks associated with frontier AI training.

OpenAI Blog
Jul 26, 2023

Frontier Model Forum

We’re forming a new industry body to promote the safe and responsible development of frontier AI systems: advancing AI safety research, identifying best practices and standards, and facilitating information sharing among policymakers and industry.

Hugging Face Trending Papers
Jun 4

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios

In this work, we propose an agentic gamification framework for hazard-informed learning of robot safety policies through synthetic scenarios. We model scenario generation as an adversarial game between two agents: a Red Team that explores the space of potential failures by constructing hazardous situations, and a Blue Team that incrementally refines safety policies to prevent them.