Google Reveals Expressive Audio Tags for Gemini 3.1 Text-to-Speech

Gemini's TTS model now supports inline tags like [screams], [whispers], and [pause] for granular vocal control — a small feature with big implications for podcast generation and interactive audio.

Google AI shared prompting tips for Gemini 3.1's TTS model, revealing that developers can embed tags like [screams], [whispers], [slow], and [pause] directly in prompts to control vocal delivery. This level of granular control over synthesized speech has previously required specialized audio engineering tools.

Unlock the full briefing

Get every story in today's briefing, the full archive, and the daily AI intelligence brief.

All stories today

Full archive

Daily brief

Cancel anytime. Payments powered by Stripe.