Gemini 3.1 Flash TTS:最强表现力语音合成模型,70语言覆盖
质量评分:5 来源: https://x.com/GoogleAI/status/2044447638511383024 抓取时间: 2026-04-18
GoogleAI · 2,205 ❤️ · 290 🔁
原文:
Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet.
This launch includes audio tags! 🗣🏷 Audio tags are a seamless way to guide vocal style, pace, and delivery using natural language commands embedded directly in your text. Want a different tempo or tone? Just tag the audio to steer the AI-speech output! The model supports 70+ languages (24 of which are high-quality evaluated languages, including: Japanese, Hindi, and Arabic). Watch the audio tags in action in the demo below ↓
中文:
今天我们发布了 Gemini 3.1 Flash TTS,这是目前表现力最强、可控性最高的文本转语音模型。
此次发布包括音频标签功能!🗣🏷 音频标签是一种无缝的方式,可以用直接嵌入文本的自然语言命令来引导语音风格、节奏和表达方式。想要不同的语速或语调?只需给音频打上标签,就能控制 AI 语音输出的风格!
该模型支持 70+ 种语言(其中 24 种为高质量评估语言,包括:日语、印地语和阿拉伯语)。在下方演示中观看音频标签的实际效果 ↓## GoogleAI · 90 ❤️ · 11 🔁
原文:
Gemini 3.1 Flash TTS is rolling out in Google Vids and is available today in preview via the Gemini API and in @GoogleAIStudio.
Whether you're creating a pitch deck or recording a passion project, transform your scripts into studio-quality narration: https://t.co/MG2YIQwKb6
中文:
今天我们发布了 Gemini 3.1 Flash TTS,这是目前表现力最强、可控性最高的文本转语音模型。
此次发布包括音频标签功能!🗣🏷 音频标签是一种无缝的方式,可以用直接嵌入文本的自然语言命令来引导语音风格、节奏和表达方式。想要不同的语速或语调?只需给音频打上标签,就能控制 AI 语音输出的风格!
该模型支持 70+ 种语言(其中 24 种为高质量评估语言,包括:日语、印地语和阿拉伯语)。在下方演示中观看音频标签的实际效果 ↓
*以上内容由 AI Field Notes 自动抓取翻译,原始内容见各推文链接。*