模型详情
GPT-5.6 Terra
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic tasks where capability and cost need to be balanced, offering strong performance at roughly half the cost of Sol.
模型规格
- 上下文长度
- 1.05M
- 最大输出
- 128K
- 输入/输出模态
- 文本 / 图像
- 发布日期
- 2026-07
02
API 端点
单一 MixRoute 网关,OpenAI 兼容
-
OpenAI-compatible
/v1/chat/completionsPOST
03
阶梯价格
单位:/1M Tokens
| 档位 | 输入 /1M Tokens | 输出 /1M Tokens | 缓存读取 /1M Tokens | 缓存写入 /1M Tokens |
|---|---|---|---|---|
| standard Length ≤ 272K | $2.0000 | $12.0000 | $0.2000 | $2.5000 |
| long_context Length > 272K | $4.0000 | $18.0000 | $0.4000 | $5.0000 |
05
选型摘要
快速判断是否适合你的工作负载
What is this model good for?
Everyday coding, reasoning, and agent tasks to test first
Use GPT-5.6 Terra for everyday coding, reasoning, and agent workflows where capability and cost need to be balanced, as stated in the official description. The full specifications and capability list are not published yet, so run one representative coding or reasoning task and compare output quality against your current model before scaling.
model: "gpt-5.6-terra", then compare with your baseline results.
What should you check before using it?
Check the 272K pricing tier boundary
Before adopting GPT-5.6 Terra, check the pricing boundary at 272K input tokens: requests at or below that length are billed $2.0000 per 1M input tokens, while longer requests move to the long-context tier at $4.0000. Measure your typical request size and estimate both tiers so the projected cost reflects your real usage.
Why use it through MixRoute?
Call it on the OpenAI-compatible endpoint with its model ID
Call GPT-5.6 Terra through MixRoute's OpenAI-compatible endpoint (POST /v1/chat/completions) with the model ID gpt-5.6-terra. Because the endpoint follows the OpenAI request format, you can keep your existing OpenAI-style client, change the model string, and test the same code path against your current setup.
/v1/chat/completions with model: "gpt-5.6-terra" and rerun your integration tests.
08
Token 成本估算
基于本页价格的实时估算,非实际账单
同一前缀被重复读取的输入占比,最高 100%
单一端点,可验证的决策
使用相同的请求格式测试此模型和备用路由
从真实工作负载开始,再由质量、总成本和故障条件决定是否接入生产流量。