OpenAI
GPT-5.6 Sol
The team’s dependable heavy-lifting builder: smart enough, practical, and willing to finish.
Canonical model record
Current identity, limits, and pricing
- Status
- Current
- API model ID
gpt-5.6-sol- Context
- 1.05M
- Max output
- 128K
- API price / 1M tokens
- $5 input · $30 output
Visual prompt runs
Benchmark runs
Open each generated scene, or compare the same prompt across models.

Helm's Deep
Fortress siege scene testing scale, lighting, architecture, and cinematic atmosphere.

Hogwarts Broom Flight Simulator
Broom-flight scene testing depth, motion cues, castle scale, and fantasy mood.

Jabberwock
Dark fantasy encounter testing creature design, forest mood, and narrative staging.

Low Poly World
Stylized island build testing composition, color, and low-poly worldbuilding.

Office Life
Workplace vignette testing everyday scene logic, objects, and believable office detail.

Petri Dish
Microscopic ecosystem testing organic forms, scientific clarity, and cellular detail.

Stormwind Trebuchet Simulator
Counterweight siege simulation testing coupled mechanics, trajectory prediction, projectile cameras, and interactive tuning.

Universe Simulator
Cosmic system testing orbital structure, glowing bodies, scale, and simulation readability.

Vice City
Neon coastal city testing vehicles, architecture, atmosphere, and dense urban layout.

Yingzao Fashi Assembly
Timber assembly scene testing structure, joinery, construction order, and material clarity.
Official benchmark profile
How GPT-5.6 Sol scores beyond our visual tests.
OpenAI presents Sol as the flagship GPT-5.6 model. These rows use the shared GPT-5.6 comparison table so the scores can be read directly against Terra, Luna, GPT-5.5, Claude, and Gemini where published.
SWE-bench Pro
64.6%Terminal-Bench 2.1
88.8%OSWorld 2.0
62.6%Full official benchmark table9 rows with source settings and peer charts
| Benchmark | Area | Score | Setting / comparison |
|---|---|---|---|
| SWE-bench Pro | Coding | 64.6% | Shared OpenAI GPT-5.6 comparison table. |
| Terminal-Bench 2.1 | Agentic coding | 88.8% | Shared OpenAI GPT-5.6 comparison table. |
| OSWorld 2.0 | Computer use | 62.6% | High-effort computer-use comparison table. |
| BrowseComp | Tool use | 90.4% | High-effort score; OpenAI separately reports Sol Ultra at 92.2%. |
| BenchCAD | Computer-aided design | 70.6% | Vision2Code score without Python tool. |
| BenchCAD with Python tool | Tool use | 83.4% | Vision2Code score with Python tool. |
| GPQA Diamond | Academic reasoning | 94.6% | Shared reasoning benchmark table. |
| FrontierMath Tier 1-3 v2 | Math | 89.0% | Shared reasoning benchmark table. |
| FrontierMath Tier 4 v2 | Math | 83.0% | Shared reasoning benchmark table. |
Superbash commentary
Our take
GPT-5.6 Sol is the team’s S-tier daily choice for demanding coding and product work. It may not feel as brilliant as Fable 5 in conversation, but it reliably gets on with the job and has been used for real app builds and migrations.
Best for
Watch out
Why it is ranked here
Evidence and commentary
Superbash editorial model ranking
Takeaway: GPT-5.6 Sol is currently placed in Tier S.
The August 2026 editorial roster places GPT-5.6 Sol at rank 2.
Open source →Keep learning