Our latest voice model has improved precision and lower latency to make voice interactions more fluid, natural and precise.
Our newest audio model introduces granular audio tags that give you precise control to direct AI speech for expressive audio generation.
Gemini 3. 5 Live Translate brings near real-time, natural speech translation to Google AI Studio, Google Translate and Google Meet.
Gemini 2. 5 Pro continues to be loved by developers as the best model for coding, and 2.
Available in preview via the API, our Computer Use model is a specialized model built on Gemini 2. 5 Pro’s capabilities to power agents that can interact with user interfaces.