
Google DeepMind
· 1 min read
Gemini 3.1 Flash TTS: the next generation of expressive AI speech
Apr 15, 2026
|
Our newest audio model introduces granular audio tags that give you precise control to direct AI speech for expressive audio generation.
Vilobh Meshram
Senior Product Manager
Max Gubin
Principal Research Engineer on behalf of the Gemini team
Your browser does not support the audio element.
Listen to article
[[duration]] minutes
This content is generated by Google AI. Generative AI is experimental
Today, we’re introducing Gemini 3.1 Flash TTS, the latest text-to-speech model that delivers improved controllability, expressivity and quality — empowering developers, enterprises and everyday users to build the next generation of AI-speech applications.
Starting today, 3.1 Flash TTS is rolling out:
- For developers in preview via the Gemini API and Google AI Studio
- For enterprises in preview on Vertex AI
- For Workspace users via Google Vids
Improved speech quality and controllability
We’ve improved the overall speech quality of Gemini 3.1 Flash TTS, making it our most natural and expressive model to date. On the Artificial Analysis TTS leaderboard, a benchmark that captures thousands of blind human preferences, 3.1 Flash TTS achieved an impressive Elo score of 1,211.
Artificial Analysis has also positioned Gemini 3.1 Flash TTS within its “most attractive quadrant” for its ideal blend of high-quality speech generation and low cost. The model stands out further with native multi-speaker dialogue, support for 70+ languages, and granular creative control via natural language.
New audio tags for more expressive speech generation
3.1 Flash TTS also introduces audio tags — an intuitive way to control vocal style, pace and delivery. By embedding natural language commands directly into the text input, you can steer AI-speech output with improved levels of granularity.
Original source
This story was published by Google DeepMind. SyncAI.news shows a preview; the complete article is on the publisher's site.
Read the full story on deepmind.google


