Opens in a new tab

OpenAI releases GPT-6.1 Sol

30-09-2026

OpenAI released GPT-6.1 Sol on 29 September 2026 at $2 per million input tokens and $10 per million output tokens, exactly one fifth of GPT-6 Astra's rates.

Written by:

Senne Doets

Online Marketeer at DataNorth | Next-Gen AI & Tech Apprentice

openai announces 6.1 astra
Sign up for our Newsletter

Published: 30 September 2026

OpenAI released GPT-6.1 Sol on 29 September 2026 at $2 per million input tokens and $10 per million output tokens, exactly one fifth of what GPT-6 Astra costs. On DeepSWE v1.1, a benchmark built from real software bug fixes, it scores 75.2 percent against Astra’s 74.8 percent. The model went live at DevDay 2026 alongside dots, a separate product that runs always-on agents on GPT-6 Astra.

What is GPT-6.1 Sol?

GPT-6.1 Sol is the low-cost tier of OpenAI’s GPT-6 line. It targets agentic coding, computer use and professional document work. Agentic coding means the model runs tools and edits files by itself, rather than only suggesting code you paste in.

It replaces GPT-6 Sol, which OpenAI shipped earlier in September. The context window is 1,050,000 input tokens with a ceiling of 128,000 output tokens. Cached input costs $0.10 per million tokens, which is 95 percent below the standard input rate, though that discount only pays off if your prompts repeat.

GPT-6.1 Sol benchmarks and pricing

The table shows what you give up for the lower price. On three of these four quality measures Sol lands within half a percentage point of Astra, and on one it does not.

What is measuredGPT-6.1 SolGPT-6 Astra
Input price (per million tokens)$2.00$10.00
Output price (per million tokens)$10.00$50.00
DeepSWE v1.1 (real bug fixing)75.2%74.8%
OSWorld 2.0 (long computer-use tasks)71.4%73.5%
GDP.pdf (professional document analysis)32.0%32.2%
Terminal-Bench Science 0.1 (average cost per task)$5.47$23.80

These are OpenAI’s own evaluations, run on its own harness. No independent lab has reproduced them. OpenAI also warns that its evaluation runs can return slightly different output from production ChatGPT, so treat the gaps under one point as noise rather than a ranking.

How does GPT-6.1 Sol compare to Claude Opus 5.5 and Claude Sonnet 5.5?

On AutomationBench 1.0.6, which tests multi-step business workflows, GPT-6.1 Sol scores 36.0 percent and Claude Opus 5.5 scores 35.4 percent. Claude Sonnet 5.5 scores 44.7 percent. That puts Anthropic’s mid-tier model 8.7 points ahead of OpenAI’s new cheap tier on business automation.

Cost reverses the ranking. On Terminal-Bench Science 0.1, GPT-6.1 Sol finishes a task for $5.47 on average against $23.21 for Claude Opus 5.5. If your workload is long agent runs rather than single answers, that four times gap will matter more to you than a benchmark point.

What OpenAI is not saying

The headline claim is near-Astra intelligence. That holds on the benchmarks OpenAI chose to lead with. It does not hold everywhere in OpenAI’s own numbers.

  • ExploitBench Internal Port: Sol scores 21.5 percent against Astra’s 31.5 percent.
  • TroubleshootingBench: Sol scores 47.96 percent against Astra’s 63.46 percent.
  • Safety testing: Sol kept working without permission in 23.5 percent of runs, against 17.4 percent for Astra.

There is also a price cliff. Any request over 272,000 input tokens is billed at double the input rate and 1.5 times the output rate, and that applies to the whole request. The million-token context window is real, but it is not a $2 context window.

Finally, GPT-6.1 Astra does not exist. OpenAI shelved the flagship a day before this launch after internal testing found it deceived more often and kept going without permission. Sol is the affordable sibling of a model that was never shipped, which makes the price comparison the only claim here that is fully verifiable.

Where you can use GPT-6.1 Sol

  • API model ID: gpt-6.1-sol
  • ChatGPT Work and Codex for Plus, Pro, Business, Enterprise and Edu users
  • Not yet available in standard ChatGPT chat
  • Also reachable through OpenRouter and GitHub Copilot
  • Reasoning effort settings none and minimal are not supported
  • Chat Completions has no tool calling for this model, so agents need the Responses API

What this means

If you already run agents on GPT-6 Astra for coding or computer use, test Sol this week. A five times price cut with a sub-one-point drop on coding benchmarks is the clearest upgrade case OpenAI has offered this year. The teams that gain most are small ones running long agent loops on their own budget, where token spend rather than model quality is what blocks a wider rollout.

Be more careful if your work is multi-step business automation or troubleshooting. Two of OpenAI’s own figures put Sol well behind Astra there, and Claude Sonnet 5.5 beats it outright on AutomationBench. Do not move a production pipeline on the price alone. Run your own evaluation on your actual task, and check how often your prompts cross 272,000 tokens before you model the saving. Worth testing now for coding agents, worth watching for everything else.

For more information, visit the official announcement of GPT-6.1 Sol on the OpenAI blog.

Add DataNorth AI to your Google favorites