Gemini 3.1 Flash Live: Making audio AI more natural and reliable
Our latest voice model has improved precision and lower latency to make voice interactions more fluid, natural and precise.
Gemini 2. 5 has new capabilities in AI-powered audio dialog and generation.
Our latest voice model has improved precision and lower latency to make voice interactions more fluid, natural and precise.
Our newest audio model introduces granular audio tags that give you precise control to direct AI speech for expressive audio generation.
Gemini 3. 5 Live Translate brings near real-time, natural speech translation to Google AI Studio, Google Translate and Google Meet.
Gemini 2. 5 Pro continues to be loved by developers as the best model for coding, and 2.
Available in preview via the API, our Computer Use model is a specialized model built on Gemini 2. 5 Pro’s capabilities to power agents that can interact with user interfaces.
The Gemini app now features our most advanced music generation model Lyria 3, empowering anyone to make 30-second tracks using text or images.
Explore the latest Gemini 2. 5 model updates with enhanced performance and accuracy: Gemini 2.
We're rolling out Deep Think in the Gemini app for Google AI Ultra subscribers, and we're giving select mathematicians access to the full version of the Gemini 2. 5 Deep Think model entered into the IMO competition.
A new generation of voice models for natural human-AI interaction, now powering ChatGPT Voice.
We’re releasing a more advanced speech-to-speech model and new API capabilities including MCP server support, image input, and SIP phone calling support.
Gemini 2. 5 Flash-Lite, previously in preview, is now stable and generally available.