qwen3-vl-embeddingCompare qwen3-vl-embedding API pricing, supported endpoints, capabilities and access options on Modelsell.
Qwen tokenizer (tiktoken-compat)Scores on standardized evaluations. Higher percentages are better — and rank percentile shows
Metrics sourced fromArtificial Analysis 2026-09-30·Qwen3 Vl Embedding
This model is not included in the current benchmark snapshot.
Missing models or measurements are not zero scores.
Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.
Qwen3-VL-Embedding 适合为商品库、设计素材和视频资料建立语义索引。用户可以输入“雨天街头撑红伞的人”,寻找内容相近的照片或视频,也可以上传一张商品图,检索外观与场景相似的素材。它把不同形式的内容转换为可比较的数值向量,帮助检索系统找到相关候选。
默认情况下,文字、图片和视频分别生成向量,便于独立搜索和比较。设置 enable_fusion=true 后,同一次输入的内容会合成一个向量。例如将产品照片与“轻量通勤双肩包,适合携带笔记本电脑”的说明一起编码,搜索时就能同时考虑外观与用途。多张图片需要分别作为 image 内容传入。
向量默认是 2,560 维,也可选择 2,048、1,536、1,024、768、512 或 256 维。较小维度有利于减少索引存储与计算量,具体召回效果应使用自己的查询样例比较。建库与查询应保持同一模型和维度,改变编码设置后重新生成对应索引。
模型输出语义向量,不直接给出答案或事件时间点。需要片段级搜索时,应用先把长视频切成带时间范围的片段,分别编码,并将向量与原视频地址、起止时间一起保存。检索后再返回相应片段;需要文字解释时,可把检索结果交给问答模型。
为商品图片建库时,把图片 URL 与商品说明放入 input.contents,例如:
[{"image":"可公开访问的商品图片URL"},{"text":"米色帆布双肩包,可放入14英寸电脑"}]
在 parameters 中设置 enable_fusion=true 和 dimension=1024,保存返回向量与商品 ID。用户查询“适合上班背的浅色电脑包”时,将查询文字编码为同维度向量,再用余弦相似度搜索商品索引。
图片可用 URL 或 Base64,视频使用可公开访问的 URL;单张图片不超过 10 MB,视频文件不超过 50 MB。先用外观相似但用途不同的商品检查误召回,再决定是否补充文字说明或增加重排序。
Use the linked documentation for model-specific request parameters and examples.
All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.
Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.
| Supplier | RPM | TPM | RPD |
|---|---|---|---|
| Alibaba | Unlimited | Unlimited | Unlimited |
| official | Unlimited | Unlimited | Unlimited |
No restriction
Compare qwen3-vl-embedding API pricing, supported endpoints, capabilities and access options on Modelsell.
Create an API key with access to qwen3-vl-embedding, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.
Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.
The model catalog lists a context window of 32000 tokens. Check the selected endpoint for request limits.
Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.
