MiniMax-MCP-Tools:一站式多模态 AI 工作流
原文:MiniMax AI | 评分:5
原文 / English
MiniMax officially released MiniMax-MCP, a one-stop multimodal AI workflow solution based on the Model Context Protocol (MCP).
This MCP server enables interaction with MiniMax's powerful Text-to-Speech, image generation, and video generation APIs. It allows MCP clients like Claude Desktop, Cursor, Windsurf, and OpenAI Agents to generate speech, clone voices, generate videos, generate images, and more.
Supported Tools
| Tool | Description | |------|-------------| | text_to_audio | Convert text to audio with a given voice | | list_voices | List all available voices | | voice_clone | Clone a voice using provided audio files | | generate_video | Generate a video from a prompt | | text_to_image | Generate an image from a prompt | | query_video_generation | Query the result of video generation tasks | | music_generation | Generate a music track from a prompt and lyrics | | voice_design | Generate a voice from a prompt using preview text |
Recent Updates
- Voice Design: New
voice_designtool — create custom voices from descriptive prompts with preview audio - Video Enhancement: Added MiniMax-Hailuo-02 model with ultra-clear quality and duration/resolution controls
- Music Generation: Enhanced
music_generationtool powered by music-1.5 model - API Host Alignment: Global uses
https://api.minimax.io, Mainland China useshttps://api.minimaxi.com
Use Cases
- Automating visual data extraction from charts and documents
- Integrating AI-powered image generation directly into chat workflows
- Adding text-to-speech functionality to AI agent responses
- Building multimodal AI agents that can see, hear, and create
中文
MiniMax 正式发布 MiniMax-MCP,这是一款基于模型上下文协议(Model Context Protocol, MCP)的一站式多模态 AI 工作流解决方案。
该 MCP 服务器可与 MiniMax 强大的文本转语音、图像生成和视频生成 API 进行交互。它使 Claude Desktop、Cursor、Windsurf、OpenAI Agents 等 MCP 客户端能够生成语音、克隆声音、生成视频、生成图像等。
支持的工具
| 工具 | 说明 | |------|------| | text_to_audio | 使用指定音色将文本转换为音频 | | list_voices | 列出所有可用音色 | | voice_clone | 使用提供的音频文件克隆声音 | | generate_video | 根据提示词生成视频 | | text_to_image | 根据提示词生成图像 | | query_video_generation | 查询视频生成任务的结果 | | music_generation | 根据提示词和歌词生成音乐 | | voice_design | 使用预览文本从提示词生成音色 |
最新更新
- Voice Design(音色设计):全新
voice_design工具,通过描述性提示词创建自定义音色,并附带预览音频 - 视频增强:新增 MiniMax-Hailuo-02 模型,支持超清晰画质以及时长/分辨率控制
- 音乐生成:
music_generation工具升级,由 music-1.5 模型提供更强能力 - API Host 对齐:Global 区使用
https://api.minimax.io,中国大陆使用https://api.minimaxi.com
应用场景
- 从图表和文档中自动化提取视觉数据
- 将 AI 图像生成功能直接集成到聊天工作流中
- 为 AI 智能体回复添加文本转语音功能
- 构建能够看、听、创造的多模态 AI 智能体
*来源:MiniMax AI | 工具:MiniMax-MCP*