OpenAI
GPT-5.6 Terra
Version 5.6 benchmark runs across the shared Superbash visual prompts.
API pricing
USD per million tokens · OpenAI API Standard (≤272K input tokens)- Input
- $2
- Output
- $12
- Cached input
- $0.2
- Context window
- 1.05M
Inspect the original visual runs
Open the generated scene to test it, or compare that prompt with another model.

Helm's Deep
Run date: 2026-07-10
Fortress siege scene testing scale, lighting, architecture, and cinematic atmosphere.

Hogwarts Broom Flight Simulator
Run date: 2026-07-10
Broom-flight scene testing depth, motion cues, castle scale, and fantasy mood.

Yingzao Fashi Assembly
Run date: 2026-07-10
Specialist test · evaluated separately
Timber assembly scene testing structure, joinery, construction order, and material clarity.
Retired tests · 6 archived runs
These tests are no longer in the active suite. Their original results remain available for inspection.
How GPT-5.6 Terra compares
Three relevant peers, side by side.
- GPT-5.6 Terra This model
- Claude Opus 4.8
- GPT-5.5
- GPT-5.6 Luna
6 benchmarks · 4 models
Official benchmark profile
Published scores outside our visual tests
SWE-bench Pro
63.4%Terminal-Bench 2.1
87.4%OSWorld 2.0
50.2%Full official benchmark table9 scores and peer comparisons
| Benchmark | Area | Score | Comparison |
|---|---|---|---|
| SWE-bench Pro | Coding | 63.4% | |
| Terminal-Bench 2.1 | Agentic coding | 87.4% | |
| OSWorld 2.0 | Computer use | 50.2% | |
| BrowseComp | Tool use | 87.5% | |
| BenchCAD | Computer-aided design | 62.3% | |
| BenchCAD with Python tool | Tool use | 78.2% | |
| GPQA Diamond | Academic reasoning | 92.9% | |
| FrontierMath Tier 1-3 v2 | Math | 84.9% | |
| FrontierMath Tier 4 v2 | Math | 68.3% |
Sources, pricing details and methodology
Verified model facts
Model specifications
- Status
- Superseded in this ranking
- API model ID
gpt-5.6-terra- Context
- 1.05M
- Max output
- 128K
- License
- Not publicly verified
- Access route
- Not publicly verified
- Modalities
- Not publicly verified
- Hardware
- Not publicly verified
- OpenAI API Standard (≤272K input tokens) · USD / 1M tokens
- $2 input · $0.2 cached input · $2.5 cache write · $12 output
Above 272K input tokens, the full request costs $4 input, $0.4 cached input, $5 cache write, and $18 output per 1M tokens. Batch, Flex, Fast mode, and regional pricing differ.
Benchmark sources
OpenAI positions Terra as the capable lower-cost GPT-5.6 option. These rows use the shared GPT-5.6 comparison table so the scores line up against Sol, Luna, GPT-5.5, Claude, and Gemini where published.
GPT-5.6: Frontier intelligence that scales with your ambitionEvaluation settings
- Shared OpenAI GPT-5.6 comparison table.
- High-effort computer-use comparison table.
- High-effort browsing-agent comparison table.
- Vision2Code score without Python tool.
- Vision2Code score with Python tool.
- Shared reasoning benchmark table.