Gemini 3.1 Flash TTS is our most controllable text-to-speech model...

@GoogleDeepMind
Google DeepMind@GoogleDeepMind
16 views Apr 21, 2026 ~1 min read
Advertisement
1
Gemini 3.1 Flash TTS is our most controllable text-to-speech model yet.

With new Audio Tags, you can easily direct vocal style, delivery, and pace through text commands. 🧵
2
🔵 More natural sounding speech
🔵 Support for 70+ languages like Hindi, Japanese, and German
🔵 SynthID watermarking on all outputs
Media image
3
Here’s where to find it:

Developers: Preview via Gemini API and @GoogleAIStudio
Enterprise: Rolling out in preview on Vertex AI
Everyone: Rolling out to @Google Vids

Find out more → goo.gle/3Orfb7I
Actions
What You Can Do
  • Export as PDF or Markdown
  • Batch Export to Notion
  • Bookmark & Highlight
  • LinkedIn & Instagram Carousel Maker
Create Free Account

Includes 7-day Premium trial

Advertisement