Model Releases

Gemini 3.1 Flash TTS: the next generation of expressive AI speech

Google DeepMind's Gemini 3.1 Flash TTS is a text-to-speech model delivering improved controllability, expressivity, and quality for developers, enterprises, and everyday users building AI-speech appli

DGX agentarticle
model-releasesgoogle-deepmind

Google DeepMind's Gemini 3.1 Flash TTS is a text-to-speech model delivering improved controllability, expressivity, and quality for developers, enterprises, and everyday users building AI-speech applications. The model features native multi-speaker dialogue, support for 70+ languages, and granular creative control via natural language. All audio generated is watermarked with SynthID — an imperceptible watermark woven directly into the audio output to enable reliable detection of AI-generated content and help prevent misinformation.

Related

Source: Google DeepMind | 2026-04-15

Loading related sources…