grok-4-1-fast-non-reasoningCompare grok-4-1-fast-non-reasoning API pricing, supported endpoints, capabilities and access options on Modelsell.
Grok tokenizer (BPE)Scores on standardized evaluations. Higher percentages are better — and rank percentile shows
Metrics sourced fromArtificial Analysis 2026-09-30·Grok 4.1 Fast Non Reasoning
This model is not included in the current benchmark snapshot.
Missing models or measurements are not zero scores.
Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.
Grok 4.1 Fast Non-Reasoning 面向直接作答的交互任务,适合解释服务规则、整理用户诉求、提取查询所需的信息,以及在应用中衔接工具。它可以从一段含糊的反馈中识别用户正在问什么,判断还缺少哪个订单号或日期,再用简洁的追问推动流程。
在客服场景中,给它清楚的业务规则、可以使用的工具和每个工具的参数说明。模型可以提出查询或操作请求,应用执行后再把真实结果传回,由模型组织回复。例如先查询订单状态,再根据返回的物流信息解释下一步处理方式;不要让模型自行编造订单数据。
它更适合条件明确、步骤较短的任务,如意图分类、材料摘要和已有规则内的答复。多项规则互相冲突、需要复杂计算或长时间规划时,应增加独立校验,或选择更适合深入推理的模型。省去较长思考过程后,仍要检查答案是否符合事实与业务条件。
联网检索和业务操作都取决于应用实际提供的工具。模型输出调用参数后,应以工具执行结果为准;失败或无记录时,回复应明确当前状态,避免把计划写成已完成。
xAI 官方 API 已从 2026 年 5 月 15 日起将这一名称重定向至 Grok 4.3,并关闭推理。已有集成应结合当前渠道响应,重新检查工具参数、回复时延和典型客服案例,确认仍满足流程要求。
例如处理催发货诉求,可提供规则与工具定义后输入:
“用户说:‘上周买的东西怎么还没到?’请先判断意图并列出查询必需信息,缺少订单号时只追问订单号。获得订单查询结果后,按真实物流状态回复;工具失败时不要说已经查到。”
程序负责执行订单查询,并把返回值作为工具结果继续传入。先用缺少订单号、查无订单、物流延迟和查询失败四类样例检查流程,再用于实际对话。涉及修改订单时,按业务系统的确认规则执行,并保存操作结果。
/v1/chat/completions| Parameter | Type | Default / range | Description |
|---|---|---|---|
reasoning_effort | enum | = medium | Controls how much the model thinks before answering |
max_completion_tokens | integer | >= 1 | Maximum tokens including hidden reasoning tokens |
stop | array | — | Up to 4 strings that stop generation |
seed | integer | — | Deterministic sampling seed (best-effort) |
stream | boolean | = false | Stream tokens via Server-Sent Events |
response_format | object | — | Force JSON object or schema-conforming output |
tools | array | — | Tool / function declarations the model may call |
tool_choice | string | autononerequired | Tool-choice policy or specific tool name |
user | string | — | End-user identifier for abuse monitoring |
Replace <YOUR_API_KEY> with the API key from your token settings.
All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.
Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.
| Parameter | Type | Default / range | Description |
|---|---|---|---|
reasoning_effort | enum | = medium | Controls how much the model thinks before answering |
max_completion_tokens | integer | >= 1 | Maximum tokens including hidden reasoning tokens |
stop | array | — | Up to 4 strings that stop generation |
seed | integer | — | Deterministic sampling seed (best-effort) |
stream | boolean | = false | Stream tokens via Server-Sent Events |
response_format | object | — | Force JSON object or schema-conforming output |
tools | array | — | Tool / function declarations the model may call |
tool_choice | string | autononerequired | Tool-choice policy or specific tool name |
user | string | — | End-user identifier for abuse monitoring |
| Supplier | RPM | TPM | RPD |
|---|---|---|---|
| xAI | Unlimited | Unlimited | Unlimited |
No restriction
Compare grok-4-1-fast-non-reasoning API pricing, supported endpoints, capabilities and access options on Modelsell.
Create an API key with access to grok-4-1-fast-non-reasoning, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.
Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.
Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.
