mimo-v2.5-tts-voicecloneCompare mimo-v2.5-tts-voiceclone API pricing, supported endpoints, capabilities and access options on Modelsell.
cl100k_baseOlder GPT-3.5 familyScores on standardized evaluations. Higher percentages are better — and rank percentile shows
Metrics sourced fromArtificial Analysis 2026-09-30·MiMo V2.5 Tts Voiceclone
This model is not included in the current benchmark snapshot.
Missing models or measurements are not zero scores.
Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.
MiMo-V2.5-TTS Voiceclone 接收一段参考音频和待朗读正文,生成具有相应音色特征的新语音。适合已经有声音样本,希望继续制作课程、旁白或多段角色内容的情况,无需先进行单独的模型训练。
参考音频决定声音方向,正文决定读出什么。选择清晰、单一说话人的 MP3 或 WAV,尽量减少背景音乐、回声和其他人声。不同录音条件会影响可提取的声音信息,制作系列内容时尽量使用一致的参考样本,并用短句试听语调、发音与相似程度。
调用中,待合成正文放在 assistant 消息,user 消息可提供情绪、节奏与场景指令。参考录音经过 Base64 编码后放入 audio.voice,并加上与真实格式一致的数据前缀。编码后的字符串不能超过 10 MB。这个字段应填写实际音频数据,不能用“模仿某种声音”的文字代替录音。
可以用自然语言或音频标签调整表演风格,但本型号不提供内置音色、纯文字声音设计或唱歌模式。低延迟流式输出尚未开放,兼容的流式请求会在推理完成后一次返回结果,应按整段合成的等待方式设计体验。
怎样让多段配音更一致? 保留同一参考录音、相近指令和文本处理方式,分段试听并检查衔接;音色相似并不保证每次表演完全相同。
只有声音描述可以使用吗? Voiceclone 需要录音,只有文字描述时应选择专门的声音设计型号。
将清晰 MP3/WAV 编码为 Base64,放入 audio.voice,例如 data:audio/wav;base64,{实际数据},并检查编码后大小。
user:保留参考声音的自然感觉,语气平静,语速适中,段落之间稍作停顿。
assistant:这一节,我们先回顾上次的实验结果,再看看新的测量方法。
先用短段试听,再分段制作长内容。不要传内置 voice 名称,也不要添加唱歌模式;按整段结果保存和播放,检查人名、数字、音量与段落首尾。
Use the linked documentation for model-specific request parameters and examples.
All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.
Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.
| Supplier | RPM | TPM | RPD |
|---|---|---|---|
| XiaomiMIMO | Unlimited | Unlimited | Unlimited |
No restriction
Compare mimo-v2.5-tts-voiceclone API pricing, supported endpoints, capabilities and access options on Modelsell.
Create an API key with access to mimo-v2.5-tts-voiceclone, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.
Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.
The model catalog lists a context window of 8192 tokens. Check the selected endpoint for request limits.
The model catalog lists a maximum output of 8192 tokens. Your request settings may set a lower limit.
Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.
