models/openai/gpt-5-6-terra
O
OpenAI·active

OpenAI: GPT-5.6 Terra

OpenAI's flagship model. Long-context specialist with 1.1M window.

Overall score
4.77
/5.00 · ranked #8
Input
$2.00
per 1M tokens
Output
$12.00
per 1M tokens
Context
1.1M
tokens
Blended
$9.50
3:1 out:in ratio

Price drops, new benchmarks, model updates. Stay current on OpenAI: GPT-5.6 Terra.

One email per change. Unsubscribe anytime.

modelpicker.aipowered by live benchmark data

Scores by test

Methodology →
Structured Output
5.0
Strategic Analysis
5.0
Constrained Rewriting
4.0
Creative Problem Solving
5.0
Tool Calling
4.0
Faithfulness
5.0
Classification
4.0
Long Context
5.0
Safety Calibration
5.0
Persona Consistency
5.0
Agentic Planning
5.0
Multilingual
5.0
Tabular Data
5.0
AIME 2025
99.7
GPQA Diamond
93.3
ARC-AGI-2
83.9
SciCode
53.9
Epoch Capabilities Index (ECI)
158.8

What you need to know

GPT-5.6 Terra is a high-reasoning model optimized for complex logic and long-form data processing. Its primary differentiator is near-perfect mathematical and strategic performance, evidenced by a 99.7% score on AIME 2025 and maximum internal ratings in strategic analysis and creative problem solving. With a 1.1M context window and a 5/5 rating for long context handling, it is built for deep analysis of massive datasets.

The model maintains high reliability across most operational tasks, achieving 5/5 in structured output, faithfulness, and agentic planning. While it ranks #7 overall, it shows slight relative weakness in classification, constrained rewriting, and tool calling, where it scores 4/5. These are not failures, but they represent the only areas where the model deviates from its otherwise maximum performance profile.

At a blended cost of $11.88/MTok, this is an expensive model. The pricing reflects its position as a premium reasoning engine rather than a general-purpose utility. Developers are paying a significant premium for the precision and the massive context window, making it inefficient for simple chat or basic classification tasks.

Use this model if your application requires complex agentic planning, high-stakes mathematical reasoning, or the processing of million-token documents. Skip this model if you are running high-volume, low-complexity tasks where the high input and output costs would outweigh the marginal gains in intelligence.

Strengths — Top 3

Structured Output5.0/5.0
Strategic Analysis5.0/5.0
Creative Problem Solving5.0/5.0

Relative weaknesses — Bottom 3

Constrained Rewriting4.0/5.0
Tool Calling4.0/5.0
Classification4.0/5.0

Similar models

XMiMo-V2.5$0.2454.69ZZ.ai: GLM 5.2$2.524.85QQwen: Qwen3.6 Max Preview$4.884.85OGPT-5.2$10.944.69