Gemini 3.1 Flash TTS - 最新的文本转语音模型
| English | 中文 | |---|---| | Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet. | 今天我们发布了Gemini 3.1 Flash TTS——我们有史以来表现力最强、可控性最高的文本转语音模型。 | | This launch [excitement] includes audio tags! 🗣🏷 | 此次发布包含音频标签功能!🗣🏷 | | Audio tags [explanatory] are a seamless way to guide vocal style, pace, and delivery using natural language commands embedded directly in your text. | 音频标签是一种将自然语言命令直接嵌入文本、无缝引导语音风格、节奏和表达方式的便捷方式。 | | Want a different tempo or tone? [amazement] Just tag the audio to steer the AI-speech output! | 想要不同的语速或语调?[惊讶] 只需为音频打上标签,即可驾驭AI语音输出! | | The model supports 70+ languages (24 of which are high-quality evaluated languages, including: Japanese, Hindi, and Arabic). | 该模型支持70多种语言(其中24种为高质量评估语言,包括日语、印地语和阿拉伯语)。 | | Watch the audio tags in action in the demo below ↓ | 请观看下方演示,了解音频标签的实际效果 ↓ |
原文(英文):
Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet.
This launch [excitement] includes audio tags! 🗣🏷 Audio tags [explanatory] are a seamless way to guide vocal style, pace, and delivery using natural language commands embedded directly in your text. Want a different tempo or tone? [amazement] Just tag the audio to steer the AI-speech output!
The model supports 70+ languages (24 of which are high-quality evaluated languages, including: Japanese, Hindi, and Arabic). Watch the audio tags in action in the demo below ↓