工具与项目 3.0 · 值得看 2026-06-07 · X

Gemini 3.1 Flash TTS:最强表现力语音合成模型,70语言覆盖

Gemini 3.1 Flash TTS:最强表现力语音合成模型,70语言覆盖

回到归档

Gemini 3.1 Flash TTS:最强表现力语音合成模型,70语言覆盖

质量评分:5 来源: https://x.com/GoogleAI/status/2044447638511383024 抓取时间: 2026-04-18

GoogleAI · 2,205 ❤️ · 290 🔁

原文:

Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet.

This launch includes audio tags! 🗣🏷 Audio tags are a seamless way to guide vocal style, pace, and delivery using natural language commands embedded directly in your text. Want a different tempo or tone? Just tag the audio to steer the AI-speech output! The model supports 70+ languages (24 of which are high-quality evaluated languages, including: Japanese, Hindi, and Arabic). Watch the audio tags in action in the demo below ↓

中文:

今天我们发布了 Gemini 3.1 Flash TTS,这是目前表现力最强、可控性最高的文本转语音模型。
此次发布包括音频标签功能!🗣🏷 音频标签是一种无缝的方式,可以用直接嵌入文本的自然语言命令来引导语音风格、节奏和表达方式。想要不同的语速或语调?只需给音频打上标签,就能控制 AI 语音输出的风格!
该模型支持 70+ 种语言(其中 24 种为高质量评估语言,包括:日语、印地语和阿拉伯语)。在下方演示中观看音频标签的实际效果 ↓## GoogleAI · 90 ❤️ · 11 🔁

原文:

Gemini 3.1 Flash TTS is rolling out in Google Vids and is available today in preview via the Gemini API and in @GoogleAIStudio.

Whether you're creating a pitch deck or recording a passion project, transform your scripts into studio-quality narration: https://t.co/MG2YIQwKb6

中文:

今天我们发布了 Gemini 3.1 Flash TTS,这是目前表现力最强、可控性最高的文本转语音模型。
此次发布包括音频标签功能!🗣🏷 音频标签是一种无缝的方式,可以用直接嵌入文本的自然语言命令来引导语音风格、节奏和表达方式。想要不同的语速或语调?只需给音频打上标签,就能控制 AI 语音输出的风格!
该模型支持 70+ 种语言(其中 24 种为高质量评估语言,包括:日语、印地语和阿拉伯语)。在下方演示中观看音频标签的实际效果 ↓

*以上内容由 AI Field Notes 自动抓取翻译,原始内容见各推文链接。*