BaiLian

Qwen Audio 3.0 Tts Plus

AlibabaToken-based
Alias:qwen-audio-3.0-tts-plus
Create API Key

Compare qwen-audio-3.0-tts-plus API pricing, supported endpoints, capabilities and access options on Modelsell.

tts语音合成streaming声音复刻声音设计指令控制
Starting price
Input / Output · 1M
Context
—
Maximum input window
Modalities
→
Released
Jul 2026

Pricing by Supplier

Alibaba
阿里巴巴百炼官方
Input$20.5882/ 1M
Output$0/ 1M

Capabilities / Supported modalities

Streaming
Input
Output

Provider & data privacy

Provider
OpenAIDocs
Tokenizer
cl100k_baseOlder GPT-3.5 family
License
Proprietary (commercial)Proprietary
Data retention30 daysNot used for upstream training by default

Performance

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows

Metrics sourced fromArtificial Analysis 2026-09-30·Qwen Audio 3.0 Tts Plus

This model is not included in the current benchmark snapshot.

Missing models or measurements are not zero scores.

Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.

About Qwen Audio 3.0 Tts Plus

让旁白有更细腻的表达

Qwen-Audio-3.0-TTS-Plus 把写好的正文合成为语音,重点面向音质、自然度与表现力要求较高的内容制作。它适合有声书叙述、宣传片旁白、角色对白和品牌语音,能够根据指令调整情绪、语气、角色感、语速与音量,让同一段文案呈现不同的表达方式。

从整体风格到句内变化

可以用自然语言描述整段的表演方向,再在正文里加入情感与富语言标签。例如用 [serious] 开始安全说明,用 [excited] 转入活动介绍;[sighing] 等标签则在指定位置加入叹息等声音。情感标签影响后续文字,遇到新标签后切换,长文本被自动分句时也要重新检查衔接。

模型支持多语言和中文方言,可使用系统音色或通过声音复刻、声音设计创建音色。系统音色与复刻音色支持指令控制;方言要选择相应系统音色或对复刻音色提出要求,声音设计音色暂不支持方言。制作系列内容时,先确定音色,再逐段微调情绪,比每段同时改变所有条件更容易保持听感一致。

试听与交付

先用包含人名、数字和长句的一小段试音,检查发音、重音和停顿,再处理完整章节。可流式接收音频,也可生成完整文件;返回的文件链接有时效,应及时保存成品。

北京地域基准价为 1.4 元/万计费字符,输出音频不另收费。汉字通常计 2 个字符,SSML 标签不计入,计费字符与 Token 数不同。美元基准按 1 美元 = 6.8 元换算,最终费用应用账户分组倍率。

Use cases and prompting

纪录片旁白示例

朗读正文:“[serious]夜幕降临,森林里的另一场旅程才刚刚开始。[amazed]看,这片叶子下藏着一只刚孵化的蝴蝶。”

表演指令:“沉稳清晰的纪录片解说,前句稍慢,发现蝴蝶时带一点惊喜,避免夸张喊叫。”

情感标签和朗读正文放在哪里?

标签直接写入待合成正文;整段表演要求放在指令字段。不要把配音要求当成需要朗读的句子。

如何给整本书配音?

固定音色,按章节或自然段合成,先统一专有名词读法和标点,再检查段落之间的音量、节奏与情绪衔接。

API access

API documentation

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Rate limits

SupplierRPMTPMRPD
AlibabaUnlimitedUnlimitedUnlimited

No restriction

Frequently asked questions about qwen-audio-3.0-tts-plus

What is qwen-audio-3.0-tts-plus?

Compare qwen-audio-3.0-tts-plus API pricing, supported endpoints, capabilities and access options on Modelsell.

How do I call qwen-audio-3.0-tts-plus?

Create an API key with access to qwen-audio-3.0-tts-plus, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is qwen-audio-3.0-tts-plus priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

How should I evaluate qwen-audio-3.0-tts-plus for my project?

Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.