BaiLian

Qwen3.6 35B A3B

AlibabaToken-based
Alias:qwen3.6-35b-a3b
Create API Key

Compare qwen3.6-35b-a3b API pricing, supported endpoints, capabilities and access options on Modelsell.

textimagevideofunction_callingtoolsstructured_outputjson_modeweb_searchvisionreasoningcontext:262144
Starting price
Input / Output · 1M
Context
262.1K
Maximum input window
Max output
235.9K
Maximum tokens per response
Modalities
→

Pricing by Supplier

official
官方接口直连
Input$0.254/ 1M
Output$1.524/ 1M
Cache Read$0.254/ 1M
Alibaba
阿里巴巴百炼官方
Input$0.254/ 1M
Output$1.524/ 1M
Cache Read$0.254/ 1M

Capabilities / Supported modalities

StreamingFunction callingToolsJSON modeStructured outputReasoningVision
Input
Output

Provider & data privacy

Provider
Alibaba (Qwen)Docs
Tokenizer
Qwen tokenizer (tiktoken-compat)
License
Tongyi Qianwen LicenseOpen weights
Data retention14 daysNot used for upstream training by default

Performance

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows

Metrics sourced fromArtificial Analysis 2026-09-30·Qwen3.6 35B A3B

This model is not included in the current benchmark snapshot.

Missing models or measurements are not zero scores.

Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.

About Qwen3.6 35B A3B

结合界面与代码推进开发

Qwen3.6-35B-A3B 适合前端实现、仓库问题分析和多步开发任务。它能阅读需求与代码,也能理解图片和视频,帮助把界面截图中的现象联系到组件、状态和交互流程。例如根据设计稿提出页面结构,比较实现截图中的间距与信息层级,或结合错误日志分析某个操作为什么没有达到预期。

处理项目问题时,提供相关目录、已有组件约定和可复现步骤,比只问“帮我修好”更容易得到可执行建议。可以先让模型缩小需要查看的文件范围,再根据文件内容生成修改,随后把运行结果传回,继续调整。它输出代码和文字;文件读写、命令执行及浏览器检查由接入的工具完成。

怎样支持连续多轮任务?

模型支持工具调用,可由智能体提供检索文件、运行检查等功能。工具定义要说明参数与返回内容,并把实际结果保留在对话中,使后续判断能够依据新证据。它还具备保留历史思考上下文的能力,适合多轮迭代;应用需要维护对应会话内容,这不等于模型会自动记住其他会话或持续训练。

默认思考模式适合代码推理和方案比较,直接改写、简单问答等任务可通过参数关闭思考。不要依赖在提示词中添加 /think 或 /nothink 来切换。复杂修改应分成可检查的小步,并以运行结果、页面表现和验收条件判断是否完成。

Use cases and prompting

例如处理前端筛选故障,提供相关组件、请求代码、报错信息和操作录像后输入:

“筛选条件已变化,但列表仍显示旧结果。请沿‘筛选状态—请求参数—响应更新—列表渲染’检查现有代码,指出证据最充分的原因。先给最小修改,再说明需要验证的操作;复用已有组件与样式,不扩展到无关页面。”

如配置了文件和执行工具,让模型读取真实文件,并将检查结果继续传入;只有建议文本时,将修改与命令交由开发环境执行。百炼请求可用 enable_thinking=false 关闭思考;简单任务不必依靠文字指令强行要求省略推理。

API access

API documentation

Code samples

RequestPOST/v1/chat/completions
Example request
Parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Replace <YOUR_API_KEY> with the API key from your token settings.

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Supported parameters

Generation parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Rate limits

SupplierRPMTPMRPD
AlibabaUnlimitedUnlimitedUnlimited
officialUnlimitedUnlimitedUnlimited

No restriction

Frequently asked questions about qwen3.6-35b-a3b

What is qwen3.6-35b-a3b?

Compare qwen3.6-35b-a3b API pricing, supported endpoints, capabilities and access options on Modelsell.

How do I call qwen3.6-35b-a3b?

Create an API key with access to qwen3.6-35b-a3b, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is qwen3.6-35b-a3b priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

What is the context window of qwen3.6-35b-a3b?

The model catalog lists a context window of 262144 tokens. Check the selected endpoint for request limits.

What is the maximum output of qwen3.6-35b-a3b?

The model catalog lists a maximum output of 235929 tokens. Your request settings may set a lower limit.

How should I evaluate qwen3.6-35b-a3b for my project?

Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.