BaiLian

Qwen3 Tts Flash Realtime

AlibabaToken-based
Alias:qwen3-tts-flash-realtime
Create API Key

Compare qwen3-tts-flash-realtime API pricing, supported endpoints, capabilities and access options on Modelsell.

tts语音合成streamingrealtime预置音色
Starting price
Input / Output · 1M
Context
—
Maximum input window
Modalities
→
Released
Nov 2025

Pricing by Supplier

Alibaba
阿里巴巴百炼官方
Input$14.7059/ 1M
Output$0/ 1M

Capabilities / Supported modalities

Streaming
Input
Output

Provider & data privacy

Provider
OpenAIDocs
Tokenizer
cl100k_baseOlder GPT-3.5 family
License
Proprietary (commercial)Proprietary
Data retention30 daysNot used for upstream training by default

Performance

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows

Metrics sourced fromArtificial Analysis 2026-09-30·Qwen3 Tts Flash Realtime

This model is not included in the current benchmark snapshot.

Missing models or measurements are not zero scores.

Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.

About Qwen3 Tts Flash Realtime

让生成中的回复逐步变成声音

Qwen3-TTS-Flash-Realtime 适合语音助手、互动讲解和客服播报。应用通过 WebSocket 会话提交文字,并持续接收合成音频,可把文本模型逐步生成的回复接入语音播放。它承担文本到声音的环节,用户语音识别、问题理解与业务查询需要由前面的流程完成。

预置音色与自然语气

模型提供多种预置音色,支持多语言、中文方言及同一音色的多语言输出,会根据正文内容自适应组织语气。可以选 Cherry 做亲切提示,或用 Ethan 做清晰讲解,再用真实回复试听。需要方言时,选择明确支持该方言的音色。

输入文本应按完整短句或自然语义分段,尽量保留标点,避免在金额、日期和名称中间提交零散字符。这样的分段更方便形成可理解的节奏,也能让应用在新内容到来时连续播放。合成前仍需处理正文中的链接、代码块和机器字段。

管理完整的实时过程

会话开始时设置音色与输出要求,分段发送正文后正确结束输入,继续接收音频直到本轮合成完成。播放器需要处理片段顺序、缓冲与用户打断;只建立 WebSocket 连接不能代表音频已经生成或播放完毕。用一整轮回复测试接收、播放和结束行为,再接入持续对话。

北京地域基准价为 1 元/万计费字符,输出音频不另收费。汉字通常计 2 个字符,SSML 标签不计入,按返回字符数结算。美元基准按 1 美元 = 6.8 元换算,最终费用应用账户分组倍率。

Use cases and prompting

分段播报一轮客服回复

建立会话并选择 Cherry,依次提交完整句子:“我已经查到您的订单。”、“包裹今天下午到达配送站,预计明天送达。”、“您可以稍后查看新的物流记录。”

结束文字输入后,继续接收并按顺序播放返回音频,直到本轮完成。

要逐字发送文本吗?

优先按完整短句或语义段发送,保留标点,避免把数字和名称切开。

用户打断后谁负责停止声音?

应用负责停止播放并取消或丢弃旧轮次音频,同时管理下一轮会话与回复。

API access

API documentation

qwen3-tts-flash-realtime

Use the linked documentation for model-specific request parameters and examples.

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Rate limits

SupplierRPMTPMRPD
AlibabaUnlimitedUnlimitedUnlimited

No restriction

Frequently asked questions about qwen3-tts-flash-realtime

What is qwen3-tts-flash-realtime?

Compare qwen3-tts-flash-realtime API pricing, supported endpoints, capabilities and access options on Modelsell.

How do I call qwen3-tts-flash-realtime?

Create an API key with access to qwen3-tts-flash-realtime, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is qwen3-tts-flash-realtime priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

How should I evaluate qwen3-tts-flash-realtime for my project?

Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.