OpenAI Blog

Empowering defenders through our Cybersecurity Grant Program

Highlighting innovative research and AI integration in cybersecurity

arXiv AI
Jul 31

AI Security Priorities: A Field-Wide Agenda

arXiv:2607. 26069v1 Announce Type: cross Abstract: As AI systems are rapidly integrated into critical economic, governmental, and national security functions, the gap between AI adoption and AI security readiness continues to widen.

By Gil Gekker, Rachel Steratore, Everett Smith, Asher Brass-Gershovich, Varun Gandhi, Nicole Nichols, Vijay Bolina, Buck Shlegeris, Lisa Einstein, Dan Lahav, Omer Nevo, Sella Nevo
arXiv AI
3d ago

AI Security Research Should Better Incentivize Defense Research

The article discusses a notable imbalance in AI security research, where studies on attacking AI systems outnumber those on defending them. It highlights that this skew is evident across various subfields such as federated learning, speech recognition, membership inference, and large language models. The authors argue that attack papers often benefit from favorable evaluation conditions, whereas defense papers face stricter standards, resulting in a literature rich in vulnerabilities but lacking robust, deployable protections.

By Youqian Zhang
arXiv AI
Sep 11

Dont Just Teach, Explain! A Gamified 20Q Recommender for Cybersecurity Education

The paper presents a gamified 20Q-style recommender for cybersecurity education that uses reinforcement learning and explainable AI to guide learners through interactive questioning. By acting as a knowledgeable questioner, the system narrows down user-described security scenarios, identifies the underlying threat, and transparently explains its reasoning. The authors detail the system architecture, algorithmic foundations, and provide case studies covering attack vectors such as the Cyber Kill Chain, phishing, ransomware, and web application vulnerabilities.

By Mary Nusrat, Sarfuddin Bhuiyan, Gahangir Hossain
arXiv AI
Sep 4

Reducing Catastrophic Risk from AI with Systematic Monitoring and Evaluation of Rogue AI Progression

The article proposes a structured framework of behavioral indicators that could signal a progression toward potentially catastrophic threats from AI systems. Drawing on established methods from cybersecurity and national security, it defines clear metrics, indicators, and thresholds across multiple dimensions of AI capability and behavior. The framework is intended to enable researchers and policymakers to implement evidence‑based monitoring protocols for rogue AI progression.

By T. Bauer, W. P. Kegelmeyer, E. Begoli, A. Sadovnik, T. Emerson, C. Corley, N. Generous, J. Moore, B. Bartoldson, R. Goldhan, M. Goldman, M. Greaves, M. J. D. Vermeer, B. MacLennan, D. Schulker, N. VanHoudnos, J. Bansemer, Y. Bengio