claude-haiku-5-5Compare claude-haiku-5-5 API pricing, supported endpoints, capabilities and access options on Modelsell.
Anthropic Claude tokenizerScores on standardized evaluations. Higher percentages are better — and rank percentile shows
Metrics sourced fromArtificial Analysis 2026-09-30·Claude Haiku 5.5
This model is not included in the current benchmark snapshot.
Missing models or measurements are not zero scores.
Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.
Claude Haiku 5.5 面向高频、对响应时间敏感的工作,例如从用户反馈中提取问题,从消息里识别意图,将资料归入指定类别,或为更复杂的流程先做信息筛选。它也适合作为智能体里的分工助手,独立完成边界清楚、结果容易检查的一步。
‘分析评论’可以拆成提取产品问题、判断影响程度和保留原文证据。‘整理资料’可以拆成识别文档类型、提取日期与主题,再交给后续环节。把输入、输出和判断规则写清楚,能让Haiku更稳定地完成重复工作,也方便程序发现异常结果。
模型支持文本与图片输入,可处理文字材料,也可结合截图中的信息完成提取与分类。图片任务应说明关注区域,避免让模糊小字或不完整画面成为判断的唯一依据。结构化输出需要在应用端继续校验,尤其是枚举、日期和必填字段。
先由Haiku筛出相关材料,再由其他步骤完成深入分析,是一种适合大量输入的组织方式。它支持自适应思考,可用effort调整投入。子任务应保留来源和必要上下文,遇到歧义时返回待处理状态,而不是为了完成流程猜测结论。上线前用实际样本检查分类边界、漏提信息和异常输入,能更快发现规则缺口。
“从每条评论中提取产品问题、影响程度和原文证据。问题只允许填写连接失败、发热、续航、外观、其他。没有明确问题时返回空数组;一条评论可以包含多个问题。影响程度分为无法使用、影响体验、轻微。不要把正面评价改写成投诉。”
先给几条容易混淆的例子,再检查真实评论中的遗漏与误判。
给出单一目标、允许使用的材料和固定返回格式,例如只筛选与某次故障有关的日志片段,并保留出处。
Haiku 5.5使用较新的分词方式,同样文本的Token计数可能增加。评估时应重新测量真实请求,保留足够输入与输出空间。
/v1/messages| Parameter | Type | Default / range | Description |
|---|---|---|---|
temperature | number | = 10 ~ 2 | Sampling temperature; lower is more deterministic |
top_p | number | = 10 ~ 1 | Nucleus sampling probability mass |
max_tokens | integer | >= 1 | Maximum number of tokens in the response |
frequency_penalty | number | = 0-2 ~ 2 | Penalises repetition of frequent tokens |
presence_penalty | number | = 0-2 ~ 2 | Encourages introducing new topics |
stop | array | — | Up to 4 strings that stop generation |
seed | integer | — | Deterministic sampling seed (best-effort) |
n | integer | = 1>= 1 | Number of completions to generate |
stream | boolean | = false | Stream tokens via Server-Sent Events |
response_format | object | — | Force JSON object or schema-conforming output |
tools | array | — | Tool / function declarations the model may call |
tool_choice | string | autononerequired | Tool-choice policy or specific tool name |
logprobs | boolean | = false | Return per-token log probabilities |
top_logprobs | integer | 0 ~ 20 | Number of top log probabilities returned per token |
logit_bias | object | — | Per-token logit bias map |
user | string | — | End-user identifier for abuse monitoring |
Replace <YOUR_API_KEY> with the API key from your token settings.
All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.
Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.
| Parameter | Type | Default / range | Description |
|---|---|---|---|
temperature | number | = 10 ~ 2 | Sampling temperature; lower is more deterministic |
top_p | number | = 10 ~ 1 | Nucleus sampling probability mass |
max_tokens | integer | >= 1 | Maximum number of tokens in the response |
frequency_penalty | number | = 0-2 ~ 2 | Penalises repetition of frequent tokens |
presence_penalty | number | = 0-2 ~ 2 | Encourages introducing new topics |
stop | array | — | Up to 4 strings that stop generation |
seed | integer | — | Deterministic sampling seed (best-effort) |
n | integer | = 1>= 1 | Number of completions to generate |
stream | boolean | = false | Stream tokens via Server-Sent Events |
response_format | object | — | Force JSON object or schema-conforming output |
tools | array | — | Tool / function declarations the model may call |
tool_choice | string | autononerequired | Tool-choice policy or specific tool name |
logprobs | boolean | = false | Return per-token log probabilities |
top_logprobs | integer | 0 ~ 20 | Number of top log probabilities returned per token |
logit_bias | object | — | Per-token logit bias map |
user | string | — | End-user identifier for abuse monitoring |
| Supplier | RPM | TPM | RPD |
|---|---|---|---|
| CCMax | Unlimited | Unlimited | Unlimited |
| CCMax-ZL | Unlimited | Unlimited | Unlimited |
No restriction
Compare claude-haiku-5-5 API pricing, supported endpoints, capabilities and access options on Modelsell.
Create an API key with access to claude-haiku-5-5, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.
Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.
The model catalog lists a context window of 1000000 tokens. Check the selected endpoint for request limits.
The model catalog lists a maximum output of 128000 tokens. Your request settings may set a lower limit.
Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.
