OpenAI

GPT 5.6 Terra

OpenAIToken-based
Alias:gpt-5.6-terra
Create API Key

Compare gpt-5.6-terra API pricing, supported endpoints, capabilities and access options on Modelsell.

textimagestreamingfunction_callingtoolsstructured_outputjson_modevisionweb_searchcachingreasoningcode_interpretercontext:1050000
Starting price
Input / Output · 1M
Context
1.1M
Maximum input window
Max output
128K
Maximum tokens per response
Modalities
→
Knowledge cutoff
Feb 2026
Released
Jul 2026

Pricing by Supplier

OpenAI
-40%
Openai官方接口
Input$2$1.2/ 1M
Output$12$7.2/ 1M
Cache Read$0.2$0.12/ 1M
Cache Write (5m)$2.5$1.5/ 1M
Cache Write (1h)$4$2.4/ 1M
Azure
-40%
微软云API直连,稳定可靠
Input$2$1.2/ 1M
Output$12$7.2/ 1M
Cache Read$0.2$0.12/ 1M
Cache Write (5m)$2.5$1.5/ 1M
Cache Write (1h)$4$2.4/ 1M
GPT官+AZ 混合
-40%
适合生产环境,az 和官 key 混合渠道
Input$2$1.2/ 1M
Output$12$7.2/ 1M
Cache Read$0.2$0.12/ 1M
Cache Write (5m)$2.5$1.5/ 1M
Cache Write (1h)$4$2.4/ 1M
GPT特价
-90%
适合vibe coding自用,gpt pro 号池
Input$2$0.2/ 1M
Output$12$1.2/ 1M
Cache Read$0.2$0.02/ 1M
Cache Write (5m)$2.5$0.25/ 1M
Cache Write (1h)$4$0.4/ 1M

Capabilities / Supported modalities

StreamingFunction callingToolsStructured outputVisionWeb searchPrompt cachingReasoningCode interpreter
Input
Output

Provider & data privacy

Provider
OpenAIDocs
Tokenizer
o200k_base
License
Proprietary (commercial)Proprietary
Data retention30 daysNot used for upstream training by default

Performance

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows

Metrics sourced fromArtificial Analysis 2026-09-30·GPT 5.6 Terra

This model is not included in the current benchmark snapshot.

Missing models or measurements are not zero scores.

Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.

About GPT 5.6 Terra

为日常业务搭建完整的处理流程

GPT-5.6 Terra 适合需要兼顾效果与使用成本的工作流,例如内部知识助手、客户问题整理、文档提取、代码解释和业务数据分析。它支持文字与图片输入,输出文本,也能以固定结构返回结果,方便接入已有系统。

知识库问答可以让它先检索相关材料,再针对用户问题形成简洁回答,并保留所依据的文档位置。面对表单截图、产品页面或流程图,可以把图片与文字要求一起提供,让它提取关键信息、解释结构或整理成后续操作清单。对于需要计算的结论,提供原始数据并配合代码工具更便于复查。

在 Responses API 中,它支持文件搜索、网页搜索、代码解释器以及项目与界面操作等工具。应用可以把检索、分析和业务查询连接成多步任务;函数调用则适合对接团队自己的接口。将工具返回的真实结果交给模型,再让它生成最终说明,可以让回答与实际业务状态保持一致。

推理投入可以从 none 到较高强度调整。固定字段提取或短回复可以先尝试较低投入;有多个限制条件的判断,再增加思考深度。投入使用前,可用一组真实业务样例检查遗漏、格式和回答依据,找到满足需求的设置。

Use cases and prompting

内部知识助手示例

“请使用知识库回答员工关于报销材料的问题。先找出适用制度和例外条件,再给出需要提交的材料清单。每项结论附文档名称或段落位置;制度没有覆盖的情况列为待咨询事项,不要补写规则。”

常见问题

  • 怎样让回答只依据公司资料?先用文件检索取得相关内容,在提示词中明确材料范围,并要求保留引用位置。
  • 字段抽取需要长篇思考吗?可以先使用 none 或较低投入,配合固定输出结构验证结果。
  • 能查询内部系统吗?应用需要配置对应函数和权限,并把查询返回结果传回模型。

API access

API documentation

Code samples

RequestPOST/v1/chat/completions
Example request
Parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Replace <YOUR_API_KEY> with the API key from your token settings.

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Supported parameters

Generation parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Rate limits

SupplierRPMTPMRPD
AzureUnlimitedUnlimitedUnlimited
defaultUnlimitedUnlimitedUnlimited
GPT官+AZ 混合UnlimitedUnlimitedUnlimited
GPT特价UnlimitedUnlimitedUnlimited
OpenAIUnlimitedUnlimitedUnlimited

No restriction

Frequently asked questions about gpt-5.6-terra

What is gpt-5.6-terra?

Compare gpt-5.6-terra API pricing, supported endpoints, capabilities and access options on Modelsell.

How do I call gpt-5.6-terra?

Create an API key with access to gpt-5.6-terra, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is gpt-5.6-terra priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

What is the context window of gpt-5.6-terra?

The model catalog lists a context window of 1050000 tokens. Check the selected endpoint for request limits.

What is the maximum output of gpt-5.6-terra?

The model catalog lists a maximum output of 128000 tokens. Your request settings may set a lower limit.

How should I evaluate gpt-5.6-terra for my project?

Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.