OpenAI logo

GPT-6.1 Sol

Multimodal
OpenAI

GPT-6.1 Sol is OpenAI's GPT-6.1 model delivering near-Astra performance for complex coding, computer use, and professional work at a lower cost. The API model id is gpt-6.1-sol. It accepts text and image input with text output, supports low, medium (API default), high, xhigh, and max reasoning effort (but not none or minimal), and provides a 1,050,000-token context window with up to 128,000 output tokens. Use the Responses API for tool calling; Chat Completions is supported without tool calling. Responses API tools include web and file search, image generation, code interpreter, hosted shell, apply patch, skills, computer use, MCP, and tool search; it also supports streaming, function calling, and structured outputs. Standard per-million-token pricing is $2 input, $0.10 cached input, $2.50 cache writes, and $10 output. Prompts above 272K input tokens cost 2x input/cache and 1.5x output across the full request. Batch and Flex cost 50% of Standard; Fast costs 2x. Regional processing adds 10% where available. Launch benchmarks are OpenAI-reported; effort and evaluation variants are recorded with each score.

Key Specifications

Parameters
-
Context
1.1M
Release Date
September 29, 2026
Average Score
58.5%

Timeline

Key dates in the model's history
Announcement
September 29, 2026
Last Update
September 30, 2026
Today
October 2, 2026

Technical Specifications

Parameters
-
Training Tokens
-
Knowledge Cutoff
April 30, 2026
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$2.00
Output (per 1M tokens)
$10.00
Max Input Tokens
1.1M
Max Output Tokens
128.0K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

Other Tests

Specialized benchmarks
AutomationBench v1.0.6
AutomationBench 1.0.6. OpenAI GPT-6.1 Sol launch interactive chart. Highest published score across the reported effort sweep: max effort (0.361). Other efforts: low 0.247, medium 0.317, high 0.332, xhigh 0.355. Evaluations used OpenAI research environments or API and may differ from production ChatGPT. • Self-reported
36.1%
DeepSWE 1.1
DeepSWE v1.1. OpenAI GPT-6.1 Sol launch interactive chart. Highest published score across the reported effort sweep: high effort (0.7522). Other efforts: low 0.6438, medium 0.7301, xhigh 0.719, max 0.719. Evaluations used OpenAI research environments or API and may differ from production ChatGPT. • Self-reported
75.2%
GDP.pdf
gdp.pdf. OpenAI GPT-6.1 Sol launch interactive chart. Highest published score across the reported effort sweep: high effort (0.32). Other efforts: low 0.27, medium 0.30, xhigh 0.318, max 0.31. Evaluations used OpenAI research environments or API and may differ from production ChatGPT. • Self-reported
32.0%
HealthBench
HealthBench, length-adjusted score (unadjusted 0.567). GPT-6.1 Sol system card addendum Table 7. • Self-reported
58.5%
HealthBench Consensus
HealthBench Consensus, length-adjusted score (unadjusted 0.959). GPT-6.1 Sol system card addendum Table 7. • Self-reported
96.0%
HealthBench Hard
HealthBench Hard, length-adjusted score (unadjusted 0.334). GPT-6.1 Sol system card addendum Table 7. • Self-reported
36.2%
HealthBench Professional
HealthBench Professional, length-adjusted score (unadjusted 0.672). GPT-6.1 Sol system card addendum Table 7. • Self-reported
64.2%
MentalHealthBench
MentalHealthBench overall at max reasoning effort (0.579). Acuity splits at max effort: non-acute 0.573, high acuity 0.591, emergent 0.580. GPT-6.1 Sol system card addendum Table 9. • Self-reported
57.9%
OSWorld 2.0
OSWorld 2.0 offline set, partial reward (v2026.08.08), not strict completion or the full task set. OpenAI GPT-6.1 Sol launch interactive chart. Highest published score across the reported effort sweep: max effort (0.7142). Other efforts: low 0.5896, medium 0.6684, high 0.6956, xhigh 0.6938. Evaluations used OpenAI research environments or API and may differ from production ChatGPT. • Self-reported
71.4%
Terminal-Bench-Science 0.1
Terminal-Bench Science 0.1. OpenAI GPT-6.1 Sol launch interactive chart. Highest published score across the reported effort sweep: max effort (0.5702). Other efforts: low 0.4371, medium 0.4756, high 0.5114, xhigh 0.5371. Evaluations used OpenAI research environments or API and may differ from production ChatGPT. • Self-reported
57.0%

License & Metadata

License
proprietary
Announcement Date
September 29, 2026
Last Updated
September 30, 2026

Compare GPT-6.1 Sol

All comparisons

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.