deepseek-v4-1-flash-260910Compare deepseek-v4-1-flash-260910 API pricing, supported endpoints, capabilities and access options on Modelsell.
Monday–Friday · 09:00–12:00 / 14:00–18:00
Monday–Friday · 00:00–09:00 / 12:00–14:00 / 18:00–24:00
Saturday / Sunday · All day
Time-based prices are determined at settlement. Tool fees are charged separately.
Monday–Friday · 09:00–12:00 / 14:00–18:00
Monday–Friday · 00:00–09:00 / 12:00–14:00 / 18:00–24:00
Saturday / Sunday · All day
Time-based prices are determined at settlement. Tool fees are charged separately.
Monday–Friday · 09:00–12:00 / 14:00–18:00
Monday–Friday · 00:00–09:00 / 12:00–14:00 / 18:00–24:00
Saturday / Sunday · All day
Time-based prices are determined at settlement. Tool fees are charged separately.
DeepSeek tokenizer (BPE)Scores on standardized evaluations. Higher percentages are better — and rank percentile shows
Metrics sourced fromArtificial Analysis 2026-09-30·DeepSeek V4 1 Flash 260910
This model is not included in the current benchmark snapshot.
Missing models or measurements are not zero scores.
Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.
DeepSeek V4.1 Flash 可以在文字任务中结合图片理解,适合分析技术文档、代码、页面截图和图表。研发人员可以把错误现场与实现代码一起提供,让它追踪问题;业务人员可以把图文资料整理成可查询、可复核的结构化记录。
这个版本支持深度思考,适合包含多项约束或需要分步判断的问题。例如,分析一段导入流程为何在部分输入下失败,先比较字段定义与异常样本,再推演执行分支;阅读一份带图表的业务报告,检查结论与数据是否相符。输出以文字、代码或约定的数据结构返回。
较大的上下文与输出空间便于处理长资料、较多代码及详细报告。实际使用时应先确定本次问题,再按文件、章节或图片编号组织内容。长答案最好约定结构和重点,分段检查关键结论,避免输出虽然完整却难以执行。
工具调用可以连接数据查询、计算或业务函数,结构化输出则方便生成工单、检查结果或字段提取记录。外部操作需要调用端提供工具和权限;建议要求它核对工具返回的时间范围、对象与状态,再继续分析。对于重要图表数字,可以提供原始数据辅助校验。
能同时看截图和代码吗? 可以将两者作为输入,并说明截图对应哪个页面、错误发生在什么操作后,让它逐项关联。
长输出应该怎样安排? 先要结论与问题索引,再按模块展开证据和修复建议;为每段保留清楚的标题和材料位置。
提供截图、相关实现和复现过程,并要求依据材料定位。
这是导入页面的报错截图、字段校验代码和三条失败样本。
请找出哪些输入触发错误,比较页面提示与后端实际规则是否一致。
按字段列出触发条件、代码位置、修复建议和回归样本。
明确区分未填写、空字符串和数值 0,不修改其他导入规则。
长文档分析先标注文档版本与优先级;结构化提取定义字段、类型和缺失值。需要执行查询或计算时,给出允许工具与验证条件,让最终答复包含证据和实际结果。
/v1/chat/completions| Parameter | Type | Default / range | Description |
|---|---|---|---|
temperature | number | = 10 ~ 2 | Sampling temperature; lower is more deterministic |
top_p | number | = 10 ~ 1 | Nucleus sampling probability mass |
max_tokens | integer | >= 1 | Maximum number of tokens in the response |
frequency_penalty | number | = 0-2 ~ 2 | Penalises repetition of frequent tokens |
presence_penalty | number | = 0-2 ~ 2 | Encourages introducing new topics |
stop | array | — | Up to 4 strings that stop generation |
seed | integer | — | Deterministic sampling seed (best-effort) |
n | integer | = 1>= 1 | Number of completions to generate |
stream | boolean | = false | Stream tokens via Server-Sent Events |
response_format | object | — | Force JSON object or schema-conforming output |
tools | array | — | Tool / function declarations the model may call |
tool_choice | string | autononerequired | Tool-choice policy or specific tool name |
logprobs | boolean | = false | Return per-token log probabilities |
top_logprobs | integer | 0 ~ 20 | Number of top log probabilities returned per token |
logit_bias | object | — | Per-token logit bias map |
user | string | — | End-user identifier for abuse monitoring |
Replace <YOUR_API_KEY> with the API key from your token settings.
All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.
Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.
| Parameter | Type | Default / range | Description |
|---|---|---|---|
temperature | number | = 10 ~ 2 | Sampling temperature; lower is more deterministic |
top_p | number | = 10 ~ 1 | Nucleus sampling probability mass |
max_tokens | integer | >= 1 | Maximum number of tokens in the response |
frequency_penalty | number | = 0-2 ~ 2 | Penalises repetition of frequent tokens |
presence_penalty | number | = 0-2 ~ 2 | Encourages introducing new topics |
stop | array | — | Up to 4 strings that stop generation |
seed | integer | — | Deterministic sampling seed (best-effort) |
n | integer | = 1>= 1 | Number of completions to generate |
stream | boolean | = false | Stream tokens via Server-Sent Events |
response_format | object | — | Force JSON object or schema-conforming output |
tools | array | — | Tool / function declarations the model may call |
tool_choice | string | autononerequired | Tool-choice policy or specific tool name |
logprobs | boolean | = false | Return per-token log probabilities |
top_logprobs | integer | 0 ~ 20 | Number of top log probabilities returned per token |
logit_bias | object | — | Per-token logit bias map |
user | string | — | End-user identifier for abuse monitoring |
| Supplier | RPM | TPM | RPD |
|---|---|---|---|
| official | Unlimited | Unlimited | Unlimited |
| Volcengine | Unlimited | Unlimited | Unlimited |
| 火山引擎特价 | Unlimited | Unlimited | Unlimited |
No restriction
Compare deepseek-v4-1-flash-260910 API pricing, supported endpoints, capabilities and access options on Modelsell.
Create an API key with access to deepseek-v4-1-flash-260910, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.
Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.
The model catalog lists a context window of 1048576 tokens. Check the selected endpoint for request limits.
The model catalog lists a maximum output of 393216 tokens. Your request settings may set a lower limit.
Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.
