BaiLian

Qwen3 Tts Instruct Flash Realtime 2026 01.22

AlibabaToken-based
Alias:qwen3-tts-instruct-flash-realtime-2026-01-22
Create API Key

Compare qwen3-tts-instruct-flash-realtime-2026-01-22 API pricing, supported endpoints, capabilities and access options on Modelsell.

tts语音合成streamingrealtime指令控制
Starting price
Input / Output · 1M
Context
—
Maximum input window
Modalities
→
Released
Jan 2026

Pricing by Supplier

Alibaba
阿里巴巴百炼官方
Input$14.7059/ 1M
Output$0/ 1M

Capabilities / Supported modalities

Streaming
Input
Output

Provider & data privacy

Provider
OpenAIDocs
Tokenizer
cl100k_baseOlder GPT-3.5 family
License
Proprietary (commercial)Proprietary
Data retention30 daysNot used for upstream training by default

Performance

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows

Metrics sourced fromArtificial Analysis 2026-09-30·Qwen3 Tts Instruct Flash Realtime 2026 01.22

This model is not included in the current benchmark snapshot.

Missing models or measurements are not zero scores.

Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.

About Qwen3 Tts Instruct Flash Realtime 2026 01.22

让实时回复表达合适的情绪

Qwen3-TTS-Instruct-Flash-Realtime-2026-01-22 通过 WebSocket 把文字生成语音,并允许用自然语言安排合成效果。它适合情境客服、互动角色、陪伴应用和需要声音表演的实时讲解。模型支持 25 种音色的中文、英文指令调节,可以描述情绪、语气和节奏,让相同内容采用不同表达。

先确定说法,再确定怎么说

把实际台词与风格指令分开。客服确认可以要求自然利落,处理用户困惑时可以要求耐心、语速稍慢;角色助手则可围绕人物性格安排活力或沉稳。指令应具体且可听出差异,不必堆砌很多形容词。先固定音色比较少量风格,再逐步调整。

流式正文仍要按完整短句组织。前面的文本模型负责生成合适回复,合成模型负责声音表达;不要让业务字段、Markdown 标记或表演说明混入待朗读台词。用户真正需要听到的内容应简短清晰,情绪表达不能遮住金额、时间等关键信息。

完整接收与播放

配置会话的音色和表演方向后,分段提交正文,正确结束输入,并持续接收本轮音频。应用需要管理缓冲、播放顺序和用户打断,停止旧轮次声音后再接入新回复。测试时同时检查语气是否合适、发音是否清楚,以及整轮播放能否正常结束。

北京地域基准价为 1 元/万计费字符,输出音频不另收费。汉字通常计 2 个字符,SSML 标签不计入,按返回字符数结算。美元基准按 1 美元 = 6.8 元换算,最终费用应用账户分组倍率。

Use cases and prompting

耐心解释一次退款进度

选择 Serena,在会话风格配置中描述:“语气温和耐心,语速稍慢,先解释进度,再清楚说明下一步。”

分段提交正文:“退款申请已经受理。”、“预计三个工作日内原路退回,到账后您会收到通知。”

结束输入后继续接收和播放,直到本轮完成。

风格指令可以用什么语言?

使用中文或英文描述,并选择支持 Instruct 调节的音色。

每句都需要更换表演要求吗?

先用统一的整体方向完成一轮回复;只有内容情境确实改变时再调整,避免语气频繁跳变。

API access

API documentation

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Rate limits

SupplierRPMTPMRPD
AlibabaUnlimitedUnlimitedUnlimited

No restriction

Frequently asked questions about qwen3-tts-instruct-flash-realtime-2026-01-22

What is qwen3-tts-instruct-flash-realtime-2026-01-22?

Compare qwen3-tts-instruct-flash-realtime-2026-01-22 API pricing, supported endpoints, capabilities and access options on Modelsell.

How do I call qwen3-tts-instruct-flash-realtime-2026-01-22?

Create an API key with access to qwen3-tts-instruct-flash-realtime-2026-01-22, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is qwen3-tts-instruct-flash-realtime-2026-01-22 priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

How should I evaluate qwen3-tts-instruct-flash-realtime-2026-01-22 for my project?

Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.