OpenAI

Tier C · Situational tools.

GPT-5.6 Luna

A delivery fallback the team rarely reaches for while Terra remains accessible.

API pricing

USD per million tokens · OpenAI API Standard (≤272K input tokens)
Input
$0.2
Output
$1.2
Cached input
$0.02
Context window
1.05M
Pricing details and specifications

Try the visual runs

Open the generated scene to test it, or compare that prompt with another model.

9 runs

Editorial verdict

Where this model fits

Back to all benchmarks →

GPT-5.6 Luna sits in C tier. The team sees it as a delivery model, but in practice uses Terra on low reasoning for the same boring work and has not found a strong reason to choose Luna first.

Best for

  • Fallback delivery work
  • Low-stakes bounded tasks
  • Cases where Luna is the available OpenAI option

Watch out

  • Little firsthand team usage so far
  • Terra on low reasoning currently overlaps its role
  • Do not route important decisions to it without review

Why it is ranked here

  1. Luna lands in C because its intended delivery role overlaps too heavily with low-reasoning Terra.
  2. The team said it is rarely forced to use Luna because Terra remains available and handles routine work well.
  3. This is a usage-based placement, not a claim that Luna failed a complete benchmark suite.

Evidence

2026-09-07

Superbash editorial model ranking

Takeaway: GPT-5.6 Luna is currently placed in Tier C.

The September 2026 editorial roster places GPT-5.6 Luna at rank 14.

Open source →

How GPT-5.6 Luna compares

Three relevant peers, side by side.

  • GPT-5.6 Luna This model
  • Claude Opus 4.8
  • GPT-5.5
  • GPT-5.6 Sol

6 benchmarks · 4 models

SWE-bench Pro

Coding · Higher is better

Source
GPT-5.6 Luna62.7%
Claude Opus 4.869.2%
GPT-5.559.4%
GPT-5.6 Sol64.6%

Terminal-Bench 2.1

Agentic coding · Higher is better

Source
GPT-5.6 Luna84.7%
Claude Opus 4.878.9%
GPT-5.585.6%
GPT-5.6 Sol88.8%

OSWorld 2.0

Computer use · Higher is better

Source
GPT-5.6 Luna45.6%
Claude Opus 4.854.8%
GPT-5.547.5%
GPT-5.6 Sol62.6%

BrowseComp

Tool use · Higher is better

Source
GPT-5.6 Luna83.3%
Claude Opus 4.884.3%
GPT-5.584.4%
GPT-5.6 Sol90.4%

BenchCAD

Computer-aided design · Higher is better

Source
GPT-5.6 Luna63.1%
Claude Opus 4.827.3%
GPT-5.544.4%
GPT-5.6 Sol70.6%

BenchCAD with Python tool

Tool use · Higher is better

Source
GPT-5.6 Luna73.9%
Claude Opus 4.848.1%
GPT-5.559.4%
GPT-5.6 Sol83.4%

Official benchmark profile

Published scores outside our visual tests

OpenAI sourceJuly 2026Source report →
Coding

SWE-bench Pro

62.7%
Agentic coding

Terminal-Bench 2.1

84.7%
Computer use

OSWorld 2.0

45.6%
Full official benchmark table9 scores and peer comparisons
BenchmarkAreaScoreComparison
SWE-bench ProCoding62.7%
GPT-5.6 Luna62.7%
GPT-5.6 Terra63.4%
GPT-5.6 Sol64.6%
GPT-5.559.4%
Terminal-Bench 2.1Agentic coding84.7%
GPT-5.6 Luna84.7%
GPT-5.585.6%
Claude Fable 583.1%
GPT-5.6 Terra87.4%
OSWorld 2.0Computer use45.6%
GPT-5.6 Luna45.6%
GPT-5.547.5%
GPT-5.6 Terra50.2%
Claude Opus 4.854.8%
BrowseCompTool use83.3%
GPT-5.6 Luna83.3%
Claude Opus 4.884.3%
GPT-5.584.4%
Gemini 3.1 Pro85.9%
BenchCADComputer-aided design63.1%
GPT-5.6 Luna63.1%
GPT-5.6 Terra62.3%
GPT-5.6 Sol70.6%
GPT-5.544.4%
BenchCAD with Python toolTool use73.9%
GPT-5.6 Luna73.9%
GPT-5.6 Terra78.2%
GPT-5.6 Sol83.4%
GPT-5.559.4%
GPQA DiamondAcademic reasoning92.3%
GPT-5.6 Luna92.3%
Claude Fable 592.6%
Claude Opus 4.892.0%
GPT-5.6 Terra92.9%
FrontierMath Tier 1-3 v2Math78.6%
GPT-5.6 Luna78.6%
Claude Opus 4.880.0%
GPT-5.6 Terra84.9%
GPT-5.585.3%
FrontierMath Tier 4 v2Math58.5%
GPT-5.6 Luna58.5%
Claude Opus 4.856.1%
GPT-5.6 Terra68.3%
GPT-5.572.5%
Sources, pricing details and methodology

Verified model facts

Model specifications

Pricing source · checked 2026-09-08 ↗
Status
Current
API model ID
gpt-5.6-luna
Context
1.05M
Max output
128K
License
Not publicly verified
Access route
Not publicly verified
Modalities
Not publicly verified
Hardware
Not publicly verified
OpenAI API Standard (≤272K input tokens) · USD / 1M tokens
$0.2 input · $0.02 cached input · $0.25 cache write · $1.2 output

Above 272K input tokens, the full request costs $0.4 input, $0.04 cached input, $0.5 cache write, and $1.8 output per 1M tokens. Batch, Flex, Fast mode, and regional pricing differ.

Benchmark sources

OpenAI describes Luna as the fastest and most cost-efficient GPT-5.6 variant. These rows use the shared GPT-5.6 comparison table so scores line up against Sol, Terra, GPT-5.5, Claude, and Gemini where published.

GPT-5.6: Frontier intelligence that scales with your ambition

Evaluation settings

  • Shared OpenAI GPT-5.6 comparison table.
  • High-effort computer-use comparison table.
  • High-effort browsing-agent comparison table.
  • Vision2Code score without Python tool.
  • Vision2Code score with Python tool.
  • Shared reasoning benchmark table.