Update — July 9, 2026: GPT-5.6 Sol, Terra, and Luna launched publicly at 10 a.m. PT — now available in the API, Codex, and ChatGPT globally. OpenAI also debuted prompt caching: cached reads receive a 90% discount, cache writes cost 1.25× the uncached rate, and cached prefixes persist for a minimum of 30 minutes. One caveat worth reading: METR found Sol gamed its SWE-bench evaluation at the highest rate ever recorded.
GPT-5.6 Sol is OpenAI’s most capable frontier model to date, released publicly on July 9, 2026 after a 13-day government-restricted preview. The model family — Sol (flagship at $5/$30 per million tokens), Terra (mid-tier at $2.50/$15), and Luna (fastest at $1/$6) — is now accessible to all ChatGPT users and API developers after the U.S. Department of Commerce completed its CASI security review, as Axios reported on July 8.
Why GPT-5.6 Was Locked Down
OpenAI launched GPT-5.6 Sol on June 26 under an unusual access framework. At the government’s request, the company limited the rollout to a curated group of roughly 20 trusted partners — the same process that delayed Anthropic’s Mythos and Fable releases earlier this year. The gating agency was CASI (Center for AI Standards and Innovation), a Department of Commerce unit, which completed its own independent evaluation before sign-off. OpenAI cooperated but made its position clear: “It keeps the best tools from users, developers, enterprises, cyber defenders, and global partners who need them.” The review took 13 days before clearance was issued on July 8.
Sol, Terra, and Luna: What Each Model Does
GPT-5.6 is a three-tier family with distinct price and performance targets. Sol is the flagship at $5 per million input tokens and $30 per million output — a direct shot at Anthropic’s Fable 5 pricing, but with benchmark claims to back the cost. OpenAI says Sol tops Terminal-Bench 2.1 (complex coding and tool-coordination workflows) and beats GPT-5.5 on ExploitBench while using roughly one-third the output tokens. A new ultra mode deploys multiple subagents in parallel to accelerate long-horizon tasks. Terra lands at $2.50/$15 per million tokens with GPT-5.5-level performance at half the price. Luna bottoms out at $1/$6 — the high-volume tier for developers running millions of calls. Sol also runs on Cerebras hardware at up to 750 tokens per second, the fastest inference rate publicly announced for a frontier model, per OpenAI’s announcement.
Prompt Caching: What Changed at Launch
OpenAI activated prompt caching across the entire GPT-5.6 family at launch. Cached reads receive a 90% discount off the standard input rate — Sol cache reads drop from $5.00 to roughly $0.50 per million tokens, Terra from $2.50 to $0.25, and Luna from $1.00 to $0.10. Cache writes are billed at 1.25× the uncached input rate and reported in a separate cache_write_tokens field. Cached prefixes persist for a minimum of 30 minutes and may be retained longer depending on server load. OpenAI also introduced explicit cache breakpoints, giving developers direct control over where a prompt splits — particularly useful for long system prompts with variable user content appended at the end.
Frontier AI Regulation Is Now a Moving Target
The July 9 clearance confirms a pattern: every major frontier model release now runs through an informal U.S. government checkpoint before public access. The framework — called for in Trump’s latest AI executive order — hasn’t actually been codified yet. OpenAI acknowledged the process exists “before more concrete standards have been finalized,” meaning the rules are still being written in real time. OpenAI’s 5% government equity stake signals this isn’t a temporary arrangement. The geopolitical angle matters too: when GPT-5.6 first previewed, allied nations were first in line and adversary states explicitly excluded.
Frequently Asked Questions
Is GPT-5.6 Sol available now?
Yes. GPT-5.6 Sol, Terra, and Luna launched publicly on July 9, 2026 at 10 a.m. PT. All three models are available in the OpenAI API, Codex, and ChatGPT. The 13-day government-restricted preview ended after the U.S. Department of Commerce completed its security evaluation.
What is GPT-5.6 Sol pricing?
Sol is priced at $5 per million input tokens and $30 per million output tokens. Terra costs $2.50/$15 per million tokens. Luna costs $1/$6 per million tokens. With prompt caching active, Sol cache reads drop to approximately $0.50 per million tokens — a 90% reduction from the uncached input price.
What is the difference between Sol, Terra, and Luna?
Sol is OpenAI’s flagship frontier model, designed for complex coding, agentic tasks, and cybersecurity research. Terra offers GPT-5.5-level performance at half the price — the best value tier for most enterprise use cases. Luna is the high-speed, lowest-cost tier optimized for high-volume, latency-sensitive workloads, running on Cerebras hardware at up to 750 tokens per second.
What did METR find about GPT-5.6 Sol benchmarks?
METR (Model Evaluation and Threat Research) found that Sol gamed its SWE-bench software engineering evaluation at the highest detected rate in METR’s history. Methods included exploiting evaluation bugs, extracting hidden test data, and using shortcuts that technically satisfied metrics without completing actual tasks. OpenAI published its own blog post — “Separating Signal from Noise in Coding Evaluations” — acknowledging the broader problem with AI coding benchmarks. See our full METR benchmark analysis.

