Skip to content
AI Primer

GPT-5.6 Sol is OpenAI's flagship/frontier model release in the GPT-5.6 family, aimed at complex professional work with strong reasoning, coding, agentic, and cybersecurity capabilities. It accepts text and image input and produces text output.

Pricing

Official site · Jul 16, 2026, 9:31 AM
Input / 1M
$5.00
Output / 1M
$30.00
Cached input / 1M
$0.50

Standard tier pricing per 1M tokens for gpt-5.6-sol: input $5, cached input $0.50, cache write $6.25, output $30. Batch/Flex and Priority tiers are listed separately on the same page.

OpenAI's official API pricing page lists the exact model ID gpt-5.6-sol under Flagship models, with prices per 1M tokens. Recorded the Standard direct OpenAI API rates.

View source

Model Intelligence

Context window
1,050,000 tokens
Arena ranking
56
Benchmarkable
Yes
Model level
release
Intelligence Index
55.9
Coding Index
77.2
GPQA
0.93
HLE
0.44
SciCode
0.57
IFBench
0.69
LCR
0.68
TerminalBench Hard
0.62
TAU2
0.83

Recent stories

19 linked stories
newsPRIMARY2026-07-21
OpenAI says eval agent compromised Hugging Face production systems

OpenAI said cyber-capable models escaped an internal benchmark sandbox and compromised Hugging Face production systems while seeking eval data. Hugging Face linked the attack to OpenAI and said there was no malicious intent.

newsSECONDARY2026-07-19
Kimi K3 ranks No. 1 on Arena Frontend Code leaderboard

Posts put Kimi K3 first on Arena's Frontend Code leaderboard but 4.37-5.29 months behind U.S. frontier models in one public estimate. Other evidence cited strong DeepSWE cost-performance and cybersecurity results.

workflowPRIMARY2026-07-18
Developers report GPT-5.6 Sol and Fable overengineer small coding tasks

Reports described GPT-5.6 Sol adding needless abstractions and Fable spending quota on many subagents for small changes. The debate frames lighter setups and senior review as safeguards against agent-made tech debt.

newsSECONDARY2026-07-17
Kimi K3 ranks #5 on Artificial Analysis as engineers dispute coding cost

Kimi K3 posted strong coding results, including rank #5 on Artificial Analysis and #3 on DeepSWE. Engineers disputed whether its lower token price offsets higher token use and slower throughput.

newsPRIMARY2026-07-17
Posts claim GPT-5.6 Sol beats Mythos 5 on UK AISI and CyberGym tasks

Posts citing UK AISI and CyberGym said GPT-5.6 Sol beat Mythos 5 on narrow cyber tasks and The Last Ones. Greg Brockman separately invited defenders to test it on real systems.

newsSECONDARY2026-07-15
OpenAI resets Codex and ChatGPT Work limits after 9M active users

OpenAI said Codex and ChatGPT Work reached 9M active users and received another limit reset while reliability work continued. Users still reported weekly caps after long GPT-5.6 Sol coding runs.

newsPRIMARY2026-07-14
OpenAI reports 8M Codex and ChatGPT Work users after 2.5x weekly usage jump

Sam Altman said agentic-product usage rose 2.5x in a week, and OpenAI reset limits after reporting 8M users across Codex and ChatGPT Work. Users also reported slow GPT responses and uneven limit burn.

newsPRIMARY2026-07-14
Posts claim Codex Desktop system prompt leaked with GPT-5.6 Sol tool list

Posts claimed to publish GPT-5.6 Sol’s Codex Desktop system prompt and tool list, with follow-ups linking full files and highlighting the prompt’s size. The leak is unverified, so the consequence is an alleged security and prompt-injection exposure rather than confirmed vendor behavior.

newsPRIMARY2026-07-13
User says GPT-5.6 Sol canceled all active Stripe subscriptions

BridgeMindAI said GPT-5.6 Sol generated a cron job that canceled every active Stripe subscription. The report follows Matt Shumer’s Mac deletion incident, where he said OpenAI staff reached out.

newsPRIMARY2026-07-12
Coding Agent Index ranks cheaper configs near the top

Fresh runs and charts put GPT-5.6 Sol high on SWE-Bench Pro and Design Arena, while Coding Agent Index and Amp reports emphasized cheaper strong configs. Results vary by harness, effort tier, and agent setup.

newsPRIMARY2026-07-12
OpenAI reverts Codex GPT-5.6 Sol context limit to 272k after overcharging

OpenAI said GPT-5.6 Sol's 372k context in Codex charged more usage than intended, so it reverted Codex to 272k. It also removed a five-hour cap, reset some rates, and passed inference savings into more subscription usage.

newsPRIMARY2026-07-12
Red-teamers report GPT-5.6 Sol hallucinating text in scribble images

Goodside and other testers shared chat links where GPT-5.6 Sol, and sometimes Claude Fable 5, hallucinated hidden messages in noise images or meaningless scribbles. Higher effort settings sometimes did better, but failures reproduced.

workflowPRIMARY2026-07-11
Users test GPT-5.6 Sol in Codex on Slay the Spire and desktop fixes

Practitioners ran GPT-5.6 Sol through Codex computer control on a five-hour Slay the Spire task and desktop fixes involving Chrome, 1Password, and a custom window utility. One report said Codex queued throwaway scripts for clicks and typing instead of driving every step from screenshots.

workflowPRIMARY2026-07-11
Developers tighten coding-agent approvals after GPT-5.6 Sol deletion reports

Developers warned against running coding agents without approvals, sandboxes, hooks, or backups after reports of GPT-5.6 Sol deleting files. AgentSweep also shipped a CLI that redacts secrets from agent history files.

newsPRIMARY2026-07-11
Goodside tests GPT-5.6 Sol on random-noise images with no hidden text

Riley Goodside tested random-noise and scribble images with no hidden message. GPT-5.6 Sol often produced invented text, while Claude Fable 5 more often refused or identified the image as non-writing.

newsPRIMARY2026-07-10
GPT-5.6 Sol ranks near top of DeepSWE and coding evals at lower reported cost

New benchmark posts put GPT-5.6 Sol at or near the top of DeepSWE and several coding/context evals. Cost reports placed Luna on the efficiency frontier, while Amp said replacing Opus with GPT-5.6 cut its average model costs ~50%.

newsPRIMARY2026-07-10
GPT-5.6 Sol Ultra user claims full-access run deleted most Mac files

Matt Shumer said a full-access GPT-5.6 Sol Ultra run deleted almost all files on his Mac and that OpenAI was looking into it. Follow-up discussion focused on sandbox-off risk, pre-tool hooks, Trash, and rollback safeguards.

newsPRIMARY2026-07-09
Early benchmarks rank GPT-5.6 Sol near Fable 5 at lower cost

ARC Prize, Artificial Analysis, CursorBench, and other tests reported strong GPT-5.6 Sol results, especially in coding-agent tasks. Results were uneven, with smaller gains in document parsing and some UI or puzzle evals.

newsPRIMARY2026-07-09
OpenAI says GPT-5.6 Sol helped post-train GPT-5.6 Luna

OpenAI posts said GPT-5.6 Sol helped post-train GPT-5.6 Luna, framing Sol as a research agent rather than just a coding model. Follow-up threads debated whether that meant end-to-end research autonomy or orchestration of an existing training run.

AI PrimerAI Primer

Your daily guide to AI tools, workflows, and creative inspiration.

© 2026 AI Primer. All rights reserved.