GPT-5.6 Sol vs GPT-5.6 Terra: Specs & Benchmark Comparison
| Characteristic | GPT-5.6 Sol | GPT-5.6 Terra |
|---|---|---|
| Company | OpenAI | OpenAI |
| Release Date | July 9, 2026 | July 9, 2026 |
| Parameters | — | — |
| Multimodal | Yes | Yes |
| Context (input) | 1.1M | 1.1M |
| Context (output) | 128K | 128K |
| Input Price / 1M | $5.00 | $2.00 |
| Output Price / 1M | $30.00 | $12.00 |
| Average Score | 0.6 | 0.6 |
| Benchmarks | ||
| ExploitBench | 0.7 | 0.5 |
| FrontierMath Tier 4 (v2) | 0.8 | 0.7 |
| Graphwalks BFS >128k | 0.9 | 0.8 |
| SEC-bench Pro | 0.7 | 0.6 |
| MedChemBench (Internal) | 0.5 | 0.3 |
| OSWorld 2.0 | 0.6 | 0.5 |
| KernelGen 1P | 0.6 | 0.5 |
| ExploitGym | 0.3 | 0.2 |
| BenchCAD | 0.7 | 0.6 |
| ARC-AGI-3 | 0.1 | 0.0 |
| FrontierCode 1.1 | 0.5 | 0.4 |
| GDP.pdf | 0.3 | 0.2 |
| Management Consulting Tasks (Internal) | 0.4 | 0.4 |
| Graphwalks BFS 1M | 0.8 | 0.7 |
| GeneBench-Pro | 0.3 | 0.2 |
| BenchCAD (with Python tool) | 0.8 | 0.8 |
| Capture-the-Flag Challenges (Internal) | 1.0 | 0.9 |
| Toolathlon | 0.6 | 0.5 |
| NanoGPT | 0.1 | 0.1 |
| FrontierMath | 0.9 | 0.8 |
| Artificial Analysis | 0.6 | 0.6 |
| LifeSciBench | 0.6 | 0.6 |
| Search and Function-Calling | 0.9 | 0.9 |
| DeepSWE | 0.7 | 0.7 |
| DeepSWE 1.1 | 0.7 | 0.7 |
| BrowseComp | 0.9 | 0.9 |
| AutomationBench | 0.2 | 0.2 |
| HealthBench Professional | 0.6 | 0.6 |
| MMMU-Pro (with tools) | 0.8 | 0.8 |
| Agents' Last Exam | 0.5 | 0.5 |
| MMMU-Pro | 0.8 | 0.8 |
| Big Finance Bench | 0.5 | 0.5 |
| MRCR v2 (8-needle) | 0.9 | 0.9 |
| GPQA | 0.9 | 0.9 |
| RSI Index | 0.6 | 0.6 |
| Terminal-Bench 2.1 | 0.9 | 0.9 |
| MRCR v2 (8-needle, 512K-1M) | 0.7 | 0.7 |
| PostTrainBench Lite | 0.5 | 0.5 |
| SWE-Bench Pro | 0.6 | 0.6 |
| Internal Research Debugging Evaluation | 0.7 | 0.7 |
| HealthBench Consensus | 1.0 | 1.0 |
| HealthBench Hard | 0.3 | 0.3 |
| Connectors | 1.0 | 1.0 |
| HealthBench | 0.6 | 0.6 |
Visual Benchmark Comparison
GPT-5.6 Sol
GPT-5.6 Terra
ExploitBench0.7 vs 0.5
0.7
0.5
FrontierMath Tier 4 (v2)0.8 vs 0.7
0.8
0.7
Graphwalks BFS >128k0.9 vs 0.8
0.9
0.8
SEC-bench Pro0.7 vs 0.6
0.7
0.6
MedChemBench (Internal)0.5 vs 0.3
0.5
0.3
OSWorld 2.00.6 vs 0.5
0.6
0.5
KernelGen 1P0.6 vs 0.5
0.6
0.5
ExploitGym0.3 vs 0.2
0.3
0.2
BenchCAD0.7 vs 0.6
0.7
0.6
ARC-AGI-30.1 vs 0.0
0.1
FrontierCode 1.10.5 vs 0.4
0.5
0.4
GDP.pdf0.3 vs 0.2
0.3
0.2
Management Consulting Tasks (Internal)0.4 vs 0.4
0.4
0.4
Graphwalks BFS 1M0.8 vs 0.7
0.8
0.7
GeneBench-Pro0.3 vs 0.2
0.3
0.2
Verdict
GPT-5.6 Terra leads in 1 out of 4 comparison categories.
Overall Performance
Both models show comparable average scores: GPT-5.6 Sol — 0.6, GPT-5.6 Terra — 0.6.
API Cost
GPT-5.6 Terra is 2.5x cheaper: input $2.00/1M vs $5.00/1M tokens.
Context Window
Same context size: 1M tokens.
Recency
Both models were released around the same time: 7/9/2026 and 7/9/2026.
More About These Models
Related Comparisons
Frequently Asked Questions
Which is better for coding — GPT-5.6 Sol or GPT-5.6 Terra?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — GPT-5.6 Sol or GPT-5.6 Terra?
GPT-5.6 Terra is cheaper for input: $2.00 per 1M tokens vs $5.00.
Which has a larger context window — GPT-5.6 Sol or GPT-5.6 Terra?
GPT-5.6 Sol supports a larger context: 1,050,000 tokens vs 1,050,000.
The GPT-5.6 Sol and GPT-5.6 Terra comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the GPT-5.6 Sol or GPT-5.6 Terra page. See also the complete list of AI model comparisons.