Mid-tier GPT-5.6 model and OpenAI's recommended replacement for o4-mini and GPT-5 mini. Built by OpenAI.
Prices updated
Input price
$2.00
per 1M tokens · Standard
Output price
$12
per 1M tokens · Standard
Input limit
1.1M
tokens
Output limit
128K
tokens
Input formats
Output formats
GPT-5.6 Terra price history
2 price records since 9/25/2026
GPT-5.6 Terra cost calculator
$100/month
Standard pricing
Overview
What is GPT-5.6 Terra?
GPT-5.6 Terra is a text & reasoning and code & slms and vision model from OpenAI. Mid-tier GPT-5.6 model and OpenAI's recommended replacement for o4-mini and GPT-5 mini. Its 1.1M context window and $2.00 input price make it a candidate for quality-focused production applications.
Benchmarks
GPT-5.6 Terra benchmarks & speed
How GPT-5.6 Terra scores on standardized evaluations, and where it lands among every model we track.
42.1
Intelligence Index
111tok/s
Output speed
1245
DesignArena Elo
Reasoning
GPQA Diamond
92.5%
Graduate-level scientific reasoning · top 12%
HLE
42.9%
Humanity's Last Exam · top 13%
Coding
SciCode
55.0%
Python for scientific computing · top 28%
Latency & design
- Time to first token
- 121.77s
- DesignArena win rate
- 48.5%
- Design battles judged
- 72,172
Independent scores from Artificial Analysis and DesignArena · updated 10/8/2026. Higher is better; ranks compare against every model we track with that score.
Rates
GPT-5.6 Terra pricing
Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.
Input · Standard
$2.00
Per 1M tokens
Cached input · Standard
$0.20
Per 1M tokens
Capabilities
GPT-5.6 Terra Tools
Tools available when using GPT-5.6 Terra through supported provider APIs.
Function calling
Supported
Structured outputs
Supported
JSON mode
Supported
Reasoning
Supported
Built-in web search
Not supported
Log probabilities
Not supported
Deterministic seed
Supported
Parallel tool calls
Not supported
Prompt caching
Supported
Where to run it
GPT-5.6 Terra API providers
7 providers serve GPT-5.6 Terra. Prices are per 1M tokens; uptime is the last 24 hours.
| Provider | Input | Output | Cached | Context | Uptime |
|---|---|---|---|---|---|
| OpenAI | $1.00 | $6.00 | $0.10 | 1.1M | 100% |
| Azure | $2.00 | $12 | $0.20 | 1.1M | 99.98% |
| OpenAI | $2.00 | $12 | $0.20 | 1.1M | 99.93% |
| Azure | $2.20 | $13 | $0.22 | 1.1M | 100% |
| Amazon Bedrock | $2.20 | $13 | $0.22 | 1.1M | 35.69% |
| Azure | $2.20 | $13 | $0.22 | 1.1M | 99.97% |
| OpenAI | $4.00 | $24 | $0.40 | 1.1M | 100% |
- Input
- $1.00
- Output
- $6.00
- Cached
- $0.10
- Context
- 1.1M
- Uptime
- 100%
- Input
- $2.00
- Output
- $12
- Cached
- $0.20
- Context
- 1.1M
- Uptime
- 99.98%
- Input
- $2.00
- Output
- $12
- Cached
- $0.20
- Context
- 1.1M
- Uptime
- 99.93%
- Input
- $2.20
- Output
- $13
- Cached
- $0.22
- Context
- 1.1M
- Uptime
- 100%
- Input
- $2.20
- Output
- $13
- Cached
- $0.22
- Context
- 1.1M
- Uptime
- 35.69%
- Input
- $2.20
- Output
- $13
- Cached
- $0.22
- Context
- 1.1M
- Uptime
- 99.97%
- Input
- $4.00
- Output
- $24
- Cached
- $0.40
- Context
- 1.1M
- Uptime
- 100%
6 of 7
Strengths and limitations
Strengths
- Measured results (AA Intelligence 42.1 · DesignArena Elo 1245 · Arena Elo 1233 · MMLU-Pro 79.1%) position it for demanding production work.
- 1.1M context supports large documents, repositories and extended conversations.
- Supports text & reasoning and code & slms and vision workloads in one model.
Limitations
- Generated tokens cost 6× more than input tokens, which matters for verbose responses.
- Large context capacity does not guarantee consistent retrieval across the entire prompt.
- Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.
Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.
Same provider
