Grok

Grok 4.1 Fast Non Reasoning

xAIToken-based
Alias:grok-4-1-fast-non-reasoning
Create API Key

Compare grok-4-1-fast-non-reasoning API pricing, supported endpoints, capabilities and access options on Modelsell.

text
Starting price
Input / Output · 1M
Context
—
Maximum input window
Modalities
→

Pricing by Supplier

xAI
-40%
xAI 官方
Input$0.2$0.12/ 1M
Output$0.5$0.3/ 1M
Cache Read$0.05$0.03/ 1M

Capabilities / Supported modalities

StreamingSystem promptFunction callingToolsJSON modeStructured outputPrompt cachingReasoning
Input
Output

Provider & data privacy

Provider
xAIDocs
Tokenizer
Grok tokenizer (BPE)
License
Proprietary (commercial)Proprietary
Data retention87 daysNot used for upstream training by default

Performance

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows

Metrics sourced fromArtificial Analysis 2026-09-30·Grok 4.1 Fast Non Reasoning

This model is not included in the current benchmark snapshot.

Missing models or measurements are not zero scores.

Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.

About Grok 4.1 Fast Non Reasoning

让客服流程及时进入下一步

Grok 4.1 Fast Non-Reasoning 面向直接作答的交互任务,适合解释服务规则、整理用户诉求、提取查询所需的信息,以及在应用中衔接工具。它可以从一段含糊的反馈中识别用户正在问什么,判断还缺少哪个订单号或日期,再用简洁的追问推动流程。

在客服场景中,给它清楚的业务规则、可以使用的工具和每个工具的参数说明。模型可以提出查询或操作请求,应用执行后再把真实结果传回,由模型组织回复。例如先查询订单状态,再根据返回的物流信息解释下一步处理方式;不要让模型自行编造订单数据。

非推理版本适合哪些问题?

它更适合条件明确、步骤较短的任务,如意图分类、材料摘要和已有规则内的答复。多项规则互相冲突、需要复杂计算或长时间规划时,应增加独立校验,或选择更适合深入推理的模型。省去较长思考过程后,仍要检查答案是否符合事实与业务条件。

能直接搜索或修改业务数据吗?

联网检索和业务操作都取决于应用实际提供的工具。模型输出调用参数后,应以工具执行结果为准;失败或无记录时,回复应明确当前状态,避免把计划写成已完成。

xAI 官方 API 已从 2026 年 5 月 15 日起将这一名称重定向至 Grok 4.3,并关闭推理。已有集成应结合当前渠道响应,重新检查工具参数、回复时延和典型客服案例,确认仍满足流程要求。

Use cases and prompting

例如处理催发货诉求,可提供规则与工具定义后输入:

“用户说:‘上周买的东西怎么还没到?’请先判断意图并列出查询必需信息,缺少订单号时只追问订单号。获得订单查询结果后,按真实物流状态回复;工具失败时不要说已经查到。”

程序负责执行订单查询,并把返回值作为工具结果继续传入。先用缺少订单号、查无订单、物流延迟和查询失败四类样例检查流程,再用于实际对话。涉及修改订单时,按业务系统的确认规则执行,并保存操作结果。

API access

API documentation

Code samples

RequestPOST/v1/chat/completions
Example request
Parameters
ParameterTypeDefault / rangeDescription
reasoning_effort
enum
=medium
Controls how much the model thinks before answering
max_completion_tokens
integer>= 1Maximum tokens including hidden reasoning tokens
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
user
string—End-user identifier for abuse monitoring

Replace <YOUR_API_KEY> with the API key from your token settings.

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Supported parameters

Generation parameters
ParameterTypeDefault / rangeDescription
reasoning_effort
enum
=medium
Controls how much the model thinks before answering
max_completion_tokens
integer>= 1Maximum tokens including hidden reasoning tokens
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
user
string—End-user identifier for abuse monitoring

Rate limits

SupplierRPMTPMRPD
xAIUnlimitedUnlimitedUnlimited

No restriction

Frequently asked questions about grok-4-1-fast-non-reasoning

What is grok-4-1-fast-non-reasoning?

Compare grok-4-1-fast-non-reasoning API pricing, supported endpoints, capabilities and access options on Modelsell.

How do I call grok-4-1-fast-non-reasoning?

Create an API key with access to grok-4-1-fast-non-reasoning, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is grok-4-1-fast-non-reasoning priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

How should I evaluate grok-4-1-fast-non-reasoning for my project?

Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.