OpenAI has launched GPT-6.1 Sol, a mid-tier model in its GPT-6 family. The company states it delivers results similar to the premium GPT-6 Astra model for coding, computer use, and professional tasks, but at one-fifth of the standard input and output rates.
In this article
Input tokens now cost $0.10 per million for cached data, which represents a 50% reduction compared to the previous GPT-6 Sol pricing. The model is live today in the OpenAI API, ChatGPT Work, and Codex under the ID `gpt-6.1-sol`.
Positioning in the GPT-6 Family
OpenAI currently offers three tiers. GPT-6 Astra costs $10 for input, $50 for output, and $1 for cached input per million tokens. GPT-6.1 Sol costs $2 for input, $10 for output, and $0.10 for cached input. GPT-6 Luna sits at the bottom with $0.10 input, $0.50 output, and $0.01 cached input.
The cached rate matters most for agents. These tools resend system prompts, tool schemas, and history on every step. Cached reads now cost 5% of the uncached input rate, down from 10% on the earlier GPT-6 Sol model.
Benchmark Results
All figures below are vendor-reported in OpenAI’s launch post. OpenAI states that competitor numbers came from public reports.
- Coding: On DeepSWE v1.1, GPT-6.1 Sol matches GPT-6 Astra at roughly one-fifth of the cost. It beats GPT-6 Sol’s best score by 6.4 percentage points, at a lower reasoning effort.
- Professional work: On GDP.pdf, which tests answers over complex professional PDFs, Sol beats Claude Opus 5.5 with fallbacks. It does so at less than half the cost per task. On AutomationBench 1.0.6, Sol scores 2.2 points above Opus 5.5 at medium effort. That result comes at roughly one-third the cost.
- Computer use: On the OSWorld 2.0 offline set, Sol gains 7 points over GPT-6 Sol at maximum effort. It lands within 2.1 points of Astra at roughly one-seventh the cost per task.
- Science: On Terminal-Bench Science 0.1, Sol more than doubles GPT-6 Sol’s score at max effort. Average cost per task is $5.47, versus $23.21 for Opus 5.5 and $23.80 for Astra. Astra still leads at 68.1%, and OpenAI recommends it for the hardest research.
- Factuality: At low effort, the share of responses with a factual error falls from 11.4% to 7.7%. That is a reduction of about 32%. The eval uses deliberately difficult conversations where users had flagged earlier model errors.
API Details Developers Need
From the GPT-6.1 Sol model page:
- Context window of 1,050,000 tokens, 128,000 max output tokens, and an April 30, 2026 knowledge cutoff.
- Text and image input; text output.
reasoning.effortaccepts low, medium (default), high, xhigh, and max. The none and minimal settings are not supported.- Use the Responses API for tool calling. Chat Completions works without tool calling.
- Prompts above 272K input tokens cost 2x input and cache rates, and 1.5x output, for the full request.
- Batch and Flex are 50% cheaper. Fast mode costs 2x standard.
- US and EU data residency are supported. Fast mode is unavailable with EU residency.
- Fine-tuning is not supported.
OpenAI also plans a GPT-6.1 Sol Ultrafast option in Codex within days. It promises up to 8x faster token generation than standard speed.
How GPT-6.1 Sol Compares
| Feature | GPT-6.1 Sol | Claude Opus 5.5 | Claude Sonnet 5.5 | Gemini 3.1 Pro Preview |
|---|---|---|---|---|
| Developer / API ID | OpenAI / gpt-6.1-sol | Anthropic / claude-opus-5-5 | Anthropic / claude-sonnet-5-5 | Google / gemini-3.1-pro-preview |
| Status | Generally available | Generally available | Generally available | Preview |
| Input price (per 1M) | $2.00 | $4.00 | $2.00 | $2.00 (≤200K prompt) |
| Output price (per 1M) | $10.00 | $20.00 | $10.00 | $12.00 (≤200K prompt) |
| Cached input read (per 1M) | $0.10 | $0.20 | $0.20 | $0.20 + $4.50/1M tokens/hr storage |
| Long-prompt surcharge | >272K input: 2x input and cache, 1.5x output | None; 1M at standard rates | None; 1M at standard rates | >200K: $4 input, $18 output |
| Context window | 1,050,000 | 1M | 1M | 1,048,576 |
| Max output tokens | 128,000 | 128K | 128K | 65,536 |
| Input modalities | Text, image | Text, image | Text, image | Text, image, video, audio, PDF |
| Reasoning control | low to max (5 levels), default medium | Adaptive thinking, always on; default medium | Adaptive thinking; default high | Thinking supported |
| Batch pricing (in / out) | 50% off standard | $2 / $10 | $1 / $5 | $1 / $6 |
| Knowledge cutoff | Apr 30, 2026 | Jun 2026 | Jun 2026 | Not listed |
| Open weights | No | No | No | No |
Standard first-party API list prices, verified September 30, 2026.
Claude Sonnet 5.5 matches Sol’s $2 and $10 list price, but Sol’s cached input costs half as much. OpenAI’s benchmarks compare Sol against Opus 5.5, not Sonnet 5.5. Gemini 3.1 Pro matches Sol on input, charges $12 for output, and remains in preview. List prices are not direct cost comparisons. Anthropic notes its newer tokenizer produces roughly 30% more tokens for the same text.
What it means
Developers using agents for coding or document analysis can switch to GPT-6.1 Sol to cut costs. The cached input rate of $0.10 per million tokens makes long-running sessions cheaper. Users should note that Astra remains the choice for difficult research tasks, while Sol offers a cheaper alternative for standard professional work.
FAQ
- What is GPT-6.1 Sol? It is OpenAI’s mid-tier GPT-6 model, released September 29, 2026, for coding, computer use, and professional work.
- How much does GPT-6.1 Sol cost? Standard API pricing is $2 per million input tokens, $0.10 cached input, and $10 output.
- Can I self-host GPT-6.1 Sol? No. It is available only through the OpenAI API, ChatGPT Work, and Codex.




