Gemini 3.1 Flash Live: Making audio AI more natural and reliable
Our latest voice model has improved precision and lower latency to make voice interactions more fluid, natural and precise.
Gemini 3. 5 Live Translate brings near real-time, natural speech translation to Google AI Studio, Google Translate and Google Meet.
Our latest voice model has improved precision and lower latency to make voice interactions more fluid, natural and precise.
Our newest audio model introduces granular audio tags that give you precise control to direct AI speech for expressive audio generation.
Gemini 2. 5 has new capabilities in AI-powered audio dialog and generation.
We're rolling out Deep Think in the Gemini app for Google AI Ultra subscribers, and we're giving select mathematicians access to the full version of the Gemini 2. 5 Deep Think model entered into the IMO competition.
The Gemini app now features our most advanced music generation model Lyria 3, empowering anyone to make 30-second tracks using text or images.
Gemini 2. 5 Pro continues to be loved by developers as the best model for coding, and 2.
Explore the latest Gemini 2. 5 model updates with enhanced performance and accuracy: Gemini 2.
Gemini 3. 5 is built to help you execute complex, agentic workflows.
We’re extending Gemini to become a world model that can make plans and imagine new experiences by simulating aspects of the world.
Explore new realtime voice models in the OpenAI API that can reason, translate, and transcribe speech, enabling more natural and intelligent voice experiences.
Developers can now build fast speech-to-speech experiences into their applications