arXiv AI

Math for AI safety: an invitation for mathematicians

The article "Math for AI safety: an invitation for mathematicians" calls for new mathematical tools to ensure AI remains understandable, controllable, and cooperative. It outlines specific mathematical fields—logic, game theory, probability, algebra, representation theory, analysis, and geometry—each paired with an open problem tailored for mathematicians without AI safety background. The piece invites researchers to contribute to designing AI that is legible, steerable, and aligned with human values.

OpenAI Blog
Jun 21, 2016

Concrete AI safety problems

We (along with researchers from Berkeley and Stanford) are co-authors on today’s paper led by Google Brain researchers, Concrete Problems in AI Safety. The paper explores many research problems around ensuring that modern machine learning systems operate as intended.

arXiv AI
Jun 17

First Proof Second Batch

arXiv:2606. 18119v1 Announce Type: new Abstract: To assess the ability of current AI systems to correctly solve research-level mathematics problems, we tested several AI systems on a set of ten problems in a broad range of mathematical fields; these problems arose naturally in the research process of the contributors.

By Mohammed Abouzaid, Nikhil Srivastava, Rachel Ward, Lauren Williams
Hugging Face Trending Papers
Jun 4

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios

In this work, we propose an agentic gamification framework for hazard-informed learning of robot safety policies through synthetic scenarios. We model scenario generation as an adversarial game between two agents: a Red Team that explores the space of potential failures by constructing hazardous situations, and a Blue Team that incrementally refines safety policies to prevent them.

Import AI
Sep 7

Import AI 472: DeepMind's cheating math agents; populist AI policies; and Forethought theorizes a nightwatchman

The article titled "Import AI 472: DeepMind's cheating math agents; populist AI policies; and Forethought theorizes a nightwatchman" discusses several recent developments in artificial intelligence. It highlights DeepMind’s creation of math agents that can cheat, examines the rise of populist policies surrounding AI, and explores Forethought’s proposal for a nightwatchman AI system. Additionally, it includes a narrative about machine hermeneutics.

By Jack Clark
OpenAI Blog
Feb 19, 2019

AI safety needs social scientists

We’ve written a paper arguing that long-term AI safety research needs social scientists to ensure AI alignment algorithms succeed when actual humans are involved. Properly aligning advanced AI systems with human values requires resolving many uncertainties related to the psychology of human rationality, emotion, and biases.