OpenAI

Tier A · Strong with clear operating rules.

GPT-5.6 Terra

The practical A-tier daily driver for routine delivery work.

Canonical model record

Current identity, limits, and pricing

Provider source · checked 2026-08-14 ↗
Status
Current
API model ID
gpt-5.6-terra
Context
1.05M
Max output
128K
API price / 1M tokens
$2.5 input · $15 output

Visual prompt runs

Benchmark runs

Open each generated scene, or compare the same prompt across models.

9 runs

Superbash commentary

Our take

Back to the full tier list →

GPT-5.6 Terra is the team’s A-tier default for more mundane implementation work. It costs less than GPT-5.5, and low-reasoning Terra is useful enough that the team rarely needs to fall back to Luna.

Best for

  • Routine coding and delivery
  • Lower-cost implementation
  • Mundane follow-through after a strong plan
  • Daily work in the OpenAI stack

Watch out

  • Not the first choice for the hardest architecture
  • Practical value depends partly on current usage limits and resets
  • Do not confuse a lower reasoning setting with a replacement for review

Why it is ranked here

  1. Terra earns A tier as the practical delivery partner to S-tier Sol.
  2. The team now treats Sol and Terra as its two OpenAI daily drivers, with Terra handling the more routine work.
  3. Current reset behavior helped the value calculation, but the team warned that this can change.

Evidence and commentary

2026-08-17

Superbash editorial model ranking

Takeaway: GPT-5.6 Terra is currently placed in Tier A.

The August 2026 editorial roster places GPT-5.6 Terra at rank 6.

Open source →

Official benchmark profile

How GPT-5.6 Terra scores beyond our visual tests.

OpenAI positions Terra as the capable lower-cost GPT-5.6 option. These rows use the shared GPT-5.6 comparison table so the scores line up against Sol, Luna, GPT-5.5, Claude, and Gemini where published.

OpenAI sourceJuly 2026Source report →
Coding

SWE-bench Pro

63.4%
Agentic coding

Terminal-Bench 2.1

87.4%
Computer use

OSWorld 2.0

50.2%
Full official benchmark table9 rows with source settings and peer charts
BenchmarkAreaScoreSetting / comparison
SWE-bench ProCoding63.4%Shared OpenAI GPT-5.6 comparison table.
Claude Fable 580.0%
Claude Opus 4.869.2%
GPT-5.6 Sol64.6%
GPT-5.6 Terra63.4%
GPT-5.6 Luna62.7%
GPT-5.559.4%
Gemini 3.1 Pro54.2%
Terminal-Bench 2.1Agentic coding87.4%Shared OpenAI GPT-5.6 comparison table.
GPT-5.6 Sol88.8%
GPT-5.6 Terra87.4%
GPT-5.585.6%
GPT-5.6 Luna84.7%
Claude Fable 583.1%
Claude Opus 4.878.9%
Gemini 3.1 Pro70.7%
OSWorld 2.0Computer use50.2%High-effort computer-use comparison table.
GPT-5.6 Sol62.6%
Claude Opus 4.854.8%
GPT-5.6 Terra50.2%
GPT-5.547.5%
GPT-5.6 Luna45.6%
BrowseCompTool use87.5%High-effort browsing-agent comparison table.
GPT-5.6 Sol Ultra92.2%
GPT-5.6 Sol90.4%
Claude Mythos 588.0%
Claude Mythos Preview87.9%
GPT-5.6 Terra87.5%
Gemini 3.1 Pro85.9%
GPT-5.584.4%
Claude Opus 4.884.3%
BenchCADComputer-aided design62.3%Vision2Code score without Python tool.
GPT-5.6 Sol70.6%
GPT-5.6 Luna63.1%
GPT-5.6 Terra62.3%
GPT-5.544.4%
Claude Mythos 538.4%
Claude Mythos Preview35.5%
Claude Opus 4.827.3%
BenchCAD with Python toolTool use78.2%Vision2Code score with Python tool.
GPT-5.6 Sol83.4%
GPT-5.6 Terra78.2%
GPT-5.6 Luna73.9%
GPT-5.559.4%
Claude Mythos 556.7%
Claude Mythos Preview56.4%
Claude Opus 4.848.1%
GPQA DiamondAcademic reasoning92.9%Shared reasoning benchmark table.
GPT-5.6 Sol94.6%
Gemini 3.1 Pro94.3%
Claude Mythos 594.1%
GPT-5.593.6%
GPT-5.6 Terra92.9%
Claude Fable 592.6%
GPT-5.6 Luna92.3%
Claude Opus 4.892.0%
FrontierMath Tier 1-3 v2Math84.9%Shared reasoning benchmark table.
GPT-5.6 Sol89.0%
Claude Fable 587.0%
GPT-5.585.3%
GPT-5.6 Terra84.9%
Claude Opus 4.880.0%
GPT-5.6 Luna78.6%
Gemini 3.1 Pro59.6%
FrontierMath Tier 4 v2Math68.3%Shared reasoning benchmark table.
Claude Fable 587.8%
GPT-5.6 Sol83.0%
GPT-5.572.5%
GPT-5.6 Terra68.3%
GPT-5.6 Luna58.5%
Claude Opus 4.856.1%