Google announced Gemini 3.1 Flash TTS, an AI speech model with improved quality and expressiveness. The model introduces audio tags that allow users to control vocal style, pace, and delivery through natural language commands across over 70 languages. It is available in Google AI Studio, Vertex AI, and Google Vids. All generated audio is watermarked with SynthID to identify AI-generated content and prevent misinformation.
No score is assigned. Sources and their independence are shown in the citation chain below.