Skip to content
AI Primer

A GPT-5.6 model designed for cost-sensitive, high-volume workloads; it supports text output, accepts text and image input, and has configurable reasoning effort.

Pricing

Model profile · Current snapshot
Input / 1M
$0.20
Output / 1M
$1.20
Blended / 1M
$0.45
Output TPS
123
TTFT (s)
109

Model Intelligence

Context window
1,050,000 tokens
Arena ranking
38
Benchmarkable
Yes
Model level
release
Intelligence Index
37.5
Coding Index
71.4
GPQA
0.91
HLE
0.4
SciCode
0.54
LCR
0.84

Recent stories

8 linked stories
releaseSECONDARY2026-08-28
Accio open-sources 107-task CommerceAgentBench

Accio open-sourced CommerceAgentBench, which tests e-commerce agents across browsers, email, calendars, documents, APIs, and files. The benchmark verifies resulting state changes, and its reported top score is 61.7%.

newsPRIMARY2026-08-08
DeepSeek V4 Flash benchmarks claim lower DeepSWE cost than GPT-5.6 Luna

Together says two V4 Flash attempts solved more DeepSWE tasks than one GPT-5.6 Luna attempt for roughly one-third the cost. Practitioners report Flash-0731 results vary sharply by harness and pass count.

workflowPRIMARY2026-08-02
Codex users route tasks across GPT-5.6 Sol, Terra, and Luna to cut token cost

Practitioners reported better Codex multi-agent runs by raising concurrency and splitting work across Sol, Terra, and Luna. One workflow sends deploy tasks to Luna Max to preserve Sol tokens.

workflowPRIMARY2026-08-01
Sol-advisor routes Codex tasks across GPT-5.6 Sol, Luna, and Terra

Sol-advisor routes Codex tasks through GPT-5.6 Sol, Luna, and Terra, while users compare Luna Max as a lower-cost reasoning setting. Early reports say small routing tests need larger benchmarks.

newsSECONDARY2026-08-01
DeepSeek V4 Flash benchmarks show cheaper tokens but 3x SWE-Bench task cost

New tests showed DeepSeek V4 Flash as cheaper per token and faster on some serving paths. Ramp said it cost 3x more than GPT-5.6 Luna per SWE-Bench task because it used more turns.

newsPRIMARY2026-07-31
OpenAI cuts GPT-5.6 Luna pricing by 80%

OpenAI said GPT-5.6 Luna pricing fell 80%, while Terra fell 20%. Codex users recommended max reasoning for linting, tests, and dependency work, but cautioned against forcing Luna into subagent roles.

newsPRIMARY2026-07-30
OpenAI cuts GPT-5.6 Luna API prices by 80%

OpenAI said GPT-5.6 Luna is 80% cheaper and Terra is 20% cheaper, with lower usage burn in Codex and ChatGPT Work. Sol Fast adds up to 2.5x speed at 2x price, and gateways reflected the new pricing.

newsPRIMARY2026-07-09
OpenAI says GPT-5.6 Sol helped post-train GPT-5.6 Luna

OpenAI posts said GPT-5.6 Sol helped post-train GPT-5.6 Luna, framing Sol as a research agent rather than just a coding model. Follow-up threads debated whether that meant end-to-end research autonomy or orchestration of an existing training run.

AI PrimerAI Primer

Your daily guide to AI tools, workflows, and creative inspiration.

© 2026 AI Primer. All rights reserved.