Google DeepMind(@GoogleDeepMind)

Gemini 3.1 Flash TTS is our most controllable text-to-speech model yet. With new Audio Tags, you ca...

5.5内容质量
Gemini 3.1 Flash TTS is our most controllable text-to-speech model yet.

With new Audio Tags, you ca...

TL;DR · AI 摘要

Google DeepMind发布Gemini 3.1 Flash TTS,支持通过文本指令控制语音风格、语调和语速。

核心要点

  • Gemini 3.1 Flash TTS是Google目前可控性最强的TTS模型
  • 新增Audio Tags功能允许用文本命令调整语音表现
  • 用户可直接控制语音的风格、节奏和表达方式
#Google DeepMind#TTS#生成式AI#语音合成
打开原文

With new Audio Tags, you can easily direct vocal style, delivery, and pace through text commands. 🧵 https://t.co/Bq4SD8eLUN" / X

Don’t miss what’s happening

Image 1: Square profile picture
Image 1: Square profile picture

Google DeepMind

@GoogleDeepMind

Gemini 3.1 Flash TTS is our most controllable text-to-speech model yet. With new Audio Tags, you can easily direct vocal style, delivery, and pace through text commands. Image 2: 🧵

4:05 PM · Apr 15, 2026

445.3K Views

Sign up now to get your own personalized timeline!