BaiLian

Qwen Audio 3.1 Tts Flash

AlibabaToken-based
Alias:qwen-audio-3.1-tts-flash
Create API Key

Compare qwen-audio-3.1-tts-flash API pricing, supported endpoints, capabilities and access options on Modelsell.

tts语音合成streaming声音复刻声音设计指令控制
Starting price
Input / Output · 1M
Context
—
Maximum input window
Modalities
→

Pricing by Supplier

Alibaba
阿里巴巴百炼官方
Input$0.220588/ 1M
Output$1.7647/ 1M

Capabilities / Supported modalities

Streaming
Input
Output

Provider & data privacy

Provider
OpenAIDocs
Tokenizer
cl100k_baseOlder GPT-3.5 family
License
Proprietary (commercial)Proprietary
Data retention30 daysNot used for upstream training by default

Performance

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows

Metrics sourced fromArtificial Analysis 2026-09-30·Qwen Audio 3.1 Tts Flash

This model is not included in the current benchmark snapshot.

Missing models or measurements are not zero scores.

Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.

About Qwen Audio 3.1 Tts Flash

把回复及时说出来

Qwen-Audio-3.1-TTS-Flash 面向语音助手、实时对话与智能客服,把已经准备好的回复文字合成为音频。它重点优化实时合成体验,支持流式生成,适合边接收声音边播放的交互流程。模型负责语音表达,业务问答、订单查询和回复内容应先由应用或文本模型完成。

让服务语气随内容变化

可以用自然语言说明说话方式,例如语气平和、语速稍慢、重点信息清晰停顿;也可在待朗读文本中加入情感标签,让提醒、安抚和祝福采用不同表达。[serious]、[excited] 等控制类标签作用于后续文字,[sighing] 等富语言标签在当前位置加入相应声音。实际对话中应减少不必要的表演,优先保证关键信息听得清楚。

模型支持多种语言和中文方言,也支持声音复刻与声音设计。选择系统音色或先创建自定义音色后再合成;复刻场景对含噪声、混响的参考音频有更强鲁棒性,仍建议使用清晰素材。方言可通过相应系统音色或复刻音色的指令设置,声音设计音色暂不支持方言。

设计连续对话体验

每轮先生成简短明确的回复,保留完整短句再送入合成,避免在金额、日期或地址中间切开。应用还需要处理播放器缓冲、用户打断与旧音频取消,防止上一轮声音继续覆盖新回复。

北京地域基准价为输入 1.5 元/百万 Token、输出 12 元/百万 Token,按返回的 usage 分别结算。美元基准按 1 美元 = 6.8 元换算,最终费用应用账户分组倍率。

Use cases and prompting

客服确认示例

朗读正文:“您的地址已更新为新地址。包裹发出后,我会提醒您查看物流。还有其他需要修改的信息吗?”

表演指令:“耐心、自然的客服语气,中等语速,在地址确认后短暂停顿,问句柔和上扬。”

能直接理解用户的语音问题吗?

此服务的输入是待朗读文字。先完成语音识别和业务处理,再把回复交给它合成。

用户说话时如何停止播报?

由应用检测打断、停止播放并取消或丢弃旧轮次音频;仅开启流式合成不会自动完成对话管理。

API access

API documentation

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Rate limits

SupplierRPMTPMRPD
AlibabaUnlimitedUnlimitedUnlimited

No restriction

Frequently asked questions about qwen-audio-3.1-tts-flash

What is qwen-audio-3.1-tts-flash?

Compare qwen-audio-3.1-tts-flash API pricing, supported endpoints, capabilities and access options on Modelsell.

How do I call qwen-audio-3.1-tts-flash?

Create an API key with access to qwen-audio-3.1-tts-flash, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is qwen-audio-3.1-tts-flash priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

How should I evaluate qwen-audio-3.1-tts-flash for my project?

Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.