Google Reveals Expressive Audio Tags for Gemini 3.1 Text-to-Speech
Gemini's TTS model now supports inline tags like [screams], [whispers], and [pause] for granular vocal control — a small feature with big implications for podcast generation and interactive audio.
Google AI shared prompting tips for Gemini 3.1's TTS model, revealing that developers can embed tags like [screams], [whispers], [slow], and [pause] directly in prompts to control vocal delivery. This level of granular control over synthesized speech has previously required specialized audio engineering tools.
Unlock the full briefing
Get every story in today's briefing, the full archive, and the daily AI intelligence brief.
All stories today
Full archive
Daily brief
Cancel anytime. Payments powered by Stripe.