grok-4-fast-non-reasoningCompare grok-4-fast-non-reasoning API pricing, supported endpoints, capabilities and access options on Modelsell.
Grok tokenizer (BPE)Scores on standardized evaluations. Higher percentages are better — and rank percentile shows
Metrics sourced fromArtificial Analysis 2026-09-30·Grok 4 Fast Non Reasoning
This model is not included in the current benchmark snapshot.
Missing models or measurements are not zero scores.
Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.
Grok 4 Fast Non-Reasoning 面向直接作答的文字任务,适合把会议记录变成行动清单、从说明文档中提取条件,或依据现有资料回答具体问题。对于规则明确、无需复杂推导的工作,可以将输出要求限定得很清楚,让它尽快返回便于下一步处理的内容。
例如整理项目周会时,可要求只提取负责人、截止日期和已确认行动,将讨论中的建议与正式决定分开。处理较长文档时,把章节或段落标上编号,并让每条提取结果附上对应位置;即使资料很多,也能按问题回到原文核查。
明确告诉模型哪些字段只能来自材料,以及缺失值如何表示。要求它保留否定词、数值和日期,不将“计划完成”改写成“已经完成”。先做事实提取,再根据这些事实生成简短回复,可以减少写作与判断混在一起带来的偏差。
它也能参与工具使用流程,例如在应用中查询资料后生成答复。网页、X 检索或代码执行需要实际提供相应工具,最终答案应依据工具返回的信息。单独使用模型时,可直接粘贴文字材料,围绕这些材料完成整理。
如果问题需要比较多项约束、证明结论或反复检查计算,给它增加独立验证步骤,或选择推理版本。非推理模式适合缩短直接任务的响应过程,但重要事实仍需检查。
xAI 官方 API 已将此名称重定向至不启用推理的 Grok 4.3。
整理会议记录时可以输入:
“请从下方记录提取已确认行动,输出任务、负责人、截止日期和原文段落编号。未确定的负责人或日期写 null,讨论建议单独列出;不要把‘考虑’‘计划’写成已完成。最后用三句话总结需要跟进的事情。记录:{带编号的会议文本}。”
先检查事实表,再使用后面的摘要。如果原记录改动了日期或负责人,应更新材料并重新提取。用于知识问答时,将问题限定到给定章节,要求找不到依据就说明缺少信息,避免生成看似完整但无出处的答案。
/v1/chat/completions| Parameter | Type | Default / range | Description |
|---|---|---|---|
reasoning_effort | enum | = medium | Controls how much the model thinks before answering |
max_completion_tokens | integer | >= 1 | Maximum tokens including hidden reasoning tokens |
stop | array | — | Up to 4 strings that stop generation |
seed | integer | — | Deterministic sampling seed (best-effort) |
stream | boolean | = false | Stream tokens via Server-Sent Events |
response_format | object | — | Force JSON object or schema-conforming output |
tools | array | — | Tool / function declarations the model may call |
tool_choice | string | autononerequired | Tool-choice policy or specific tool name |
user | string | — | End-user identifier for abuse monitoring |
Replace <YOUR_API_KEY> with the API key from your token settings.
All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.
Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.
| Parameter | Type | Default / range | Description |
|---|---|---|---|
reasoning_effort | enum | = medium | Controls how much the model thinks before answering |
max_completion_tokens | integer | >= 1 | Maximum tokens including hidden reasoning tokens |
stop | array | — | Up to 4 strings that stop generation |
seed | integer | — | Deterministic sampling seed (best-effort) |
stream | boolean | = false | Stream tokens via Server-Sent Events |
response_format | object | — | Force JSON object or schema-conforming output |
tools | array | — | Tool / function declarations the model may call |
tool_choice | string | autononerequired | Tool-choice policy or specific tool name |
user | string | — | End-user identifier for abuse monitoring |
| Supplier | RPM | TPM | RPD |
|---|---|---|---|
| xAI | Unlimited | Unlimited | Unlimited |
No restriction
Compare grok-4-fast-non-reasoning API pricing, supported endpoints, capabilities and access options on Modelsell.
Create an API key with access to grok-4-fast-non-reasoning, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.
Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.
Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.
