GPT-5.6 Terra
The everyday tier of the GPT-5.6 family, roughly corresponding to the old 'mini' slot but benchmarking above the previous generation's flagship. For most application backends this is the default worth trying before reaching for Sol, since it clears Claude Fable 5 on several evals at a fraction of the cost.
Pricing checked against provider documentation on . How we verify
gpt-5.6-terra reasoning Input
$2.00
/ 1M tokens
Output
$12.00
/ 1M tokens
Context
1.05M
1,048,576 tokens
Max output
128K
tokens per request
Pricing caveat: Cut 20% from the $2.50/$15 launch price on 2026-07-30.
What it costs in practice
A workload of 10M input and 2M output tokens per month — roughly a moderately busy production assistant — runs $44.00/month on GPT-5.6 Terra. Routing that same volume through the batch API brings it to $22.00. The next cheaper option, Claude Sonnet 5 , would cost $40.00.
Where GPT-5.6 Terra fits
- Customer support and chat backends
- RAG and document processing pipelines
- Mid-complexity coding tasks
- Tool-calling agents at scale
Specifications
- API model ID
- gpt-5.6-terra
- Provider
- OpenAI
- Model family
- GPT-5.6
- Released
- July 2026
- Knowledge cutoff
- February 2026
- Context window
- 1,048,576 tokens
- Max output
- 128,000 tokens
- Native reasoning
- Yes
- Relative latency
- medium
- Input modalities
- text, image, audio
- Output modalities
- text
- Open weights
- No
- Cached input
- $0.20 / 1M tokens
- Batch discount
- 50%
Frequently asked
How much does GPT-5.6 Terra cost?
GPT-5.6 Terra costs $2.00 per million input tokens and $12.00 per million output tokens, with cached input at $0.20. Batch processing is 50% cheaper.
What is GPT-5.6 Terra's context window?
GPT-5.6 Terra accepts up to 1,048,576 tokens of context and can generate up to 128,000 output tokens per request.
Is GPT-5.6 Terra a reasoning model?
Yes. GPT-5.6 Terra performs native chain-of-thought before answering. Those thinking tokens are billed at the output rate, so budget above the sticker price.
What is GPT-5.6 Terra best for?
Balanced production workloads. The everyday tier of the GPT-5.6 family, roughly corresponding to the old 'mini' slot but benchmarking above the previous generation's flagship. For most application backends this is the default worth trying before reaching for Sol, since it clears Claude Fable 5 on several evals at a fraction of the cost.