qwen3-tts-instruct-flash-realtimeCompare qwen3-tts-instruct-flash-realtime API pricing, supported endpoints, capabilities and access options on Modelsell.
cl100k_baseOlder GPT-3.5 familyScores on standardized evaluations. Higher percentages are better — and rank percentile shows
Metrics sourced fromArtificial Analysis 2026-09-30·Qwen3 Tts Instruct Flash Realtime
This model is not included in the current benchmark snapshot.
Missing models or measurements are not zero scores.
Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.
Qwen3 TTS Instruct Flash Realtime 将给定文字合成为语音,并允许通过自然语言指令调整表达方式。可以选择支持的音色,再说明希望使用怎样的情绪、节奏和重音,适合客服播报、广告介绍、互动角色和讲解旁白。
它处理的是朗读表现。问答内容、事实核对和文案改写应先完成,再将确定的正文送入合成。把“温和但不拖沓”“重点强调最后一句”等要求放在独立的 instructions 中,正文只保留需要读出的内容,避免将表演要求一起读出来。
先固定音色和正文,每轮只调整一两个方面,例如语速与句末停顿。指令支持中文或英文,避免同时要求“非常快”与“每字都慢读”这类冲突。可使用 optimize_instructions 优化指令表达,再通过试听比较是否符合实际需要。
通过 WebSocket 配置会话后,可以持续提交文本并接收音频。server_commit 由服务端决定合成时机;commit 由客户端手动提交,需要自行保证句子完整。长内容按语义分段,避免把数字、姓名或一句话切在不自然的位置。结束输入后继续接收剩余音频,确认最后一句没有被截断。
北京地域基准价为 1 元 / 万计费字符,输出音频不另收费,按供应商返回的计费字符数结算。字符数不同于文本长度或 Token 数:汉字通常计 2 个字符,SSML 标签不计入。
系统内部价格以美元保存,按 1 美元 = 6.8 元换算;最终费用应用账户分组倍率。供应商账号免费额度和活动优惠不计入该基准价。
先选择支持的音色,例如 Cherry,并在会话配置中分别设置指令和朗读正文。
指令:“像讲解员一样自然清晰,中等语速;开头亲切,最后一句略放慢,不使用夸张的广告腔。” 正文:“欢迎来到夜间观星活动。请先关闭强光手电,给眼睛一点适应黑暗的时间。准备好以后,我们一起寻找北极星。”
用 session.update 设置音色、音频格式与 instructions,通过文本缓冲事件追加正文。首次可用 server_commit,持续接收音频并在输入结束时正常结束会话。试听专有名词、停顿和最后一句,再调整指令;不要把声音风格指令当作创建新音色或声音复刻。
Use the linked documentation for model-specific request parameters and examples.
All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.
Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.
| Supplier | RPM | TPM | RPD |
|---|---|---|---|
| Alibaba | Unlimited | Unlimited | Unlimited |
| official | Unlimited | Unlimited | Unlimited |
No restriction
Compare qwen3-tts-instruct-flash-realtime API pricing, supported endpoints, capabilities and access options on Modelsell.
Create an API key with access to qwen3-tts-instruct-flash-realtime, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.
Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.
Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.
