OpenAI Blog

Vallée Duhamel & Sora

Filmmaking duo Vallée Duhamel explains how Sora helps build new worlds.

OpenAI Blog
Mar 25, 2024

Sora first impressions

Since we introduced Sora to the world last month, we’ve been working with artists to learn how Sora might aid in their creative process.

OpenAI Blog
Sep 30, 2025

Sora 2 System Card

Sora 2 is our new state of the art video and audio generation model. Building on the foundation of Sora, this new model introduces capabilities that have been difficult for prior video models to achieve– such as more accurate physics, sharper realism, synchronized audio, enhanced steerability, and an expanded stylistic range.

OpenAI Blog
Sep 30, 2025

Sora 2 is here

Our latest video generation model is more physically accurate, realistic, and controllable than prior systems. It also features synchronized dialogue and sound effects.

OpenAI Blog
Dec 9, 2024

Sora System Card

Sora is OpenAI’s video generation model, designed to take text, image, and video inputs and generate a new video as an output. Sora builds on learnings from DALL-E and GPT models, and is designed to give people expanded tools for storytelling and creative expression.

OpenAI Blog
Mar 23

Creating with Sora Safely

To address the novel safety challenges posed by a state-of-the-art video model as well as a new social creation platform, we’ve built Sora 2 and the Sora app with safety at the foundation. Our approach is anchored in concrete protections.

OpenAI Blog
Sep 30, 2025

Launching Sora responsibly

To address the novel safety challenges posed by a state-of-the-art video model as well as a new social creation platform, we’ve built Sora 2 and the Sora app with safety at the foundation. Our approach is anchored in concrete protections.

Hugging Face Trending Papers
Jun 8

EditSSC: Toward Editable Semantic Occupancy Scenes with Unconditional Diffusion Models

3D semantic scene generation is crucial for autonomous driving applications, yet most methods rely on complex 3D-specific architectures such as triplane encoders and adapted diffusion networks, limiting both their simplicity and their editing capabilities. We propose EditSSC, an editing-ready method for 3D semantic scene generation using 2D Bird's Eye View (BEV) representations and off-the-shelf latent diffusion network.