Close Menu
WithO2WithO2

    Subscribe to Updates

    Get the latest AI News Tools Updates in your Inbox

    What's Hot

    Nvidia Is Financing the Companies That Buy Its Own GPUs

    July 20, 2026

    Grok Build CLI Sends Your Code Sessions to xAI Without Warning

    July 20, 2026

    Meta Iris AI Chip Hits Production in September

    July 20, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram
    Facebook X (Twitter) Instagram
    WithO2WithO2
    • AI
    • Blog
    • Business Software
    • Trending News
    • Stories
    WithO2WithO2
    Home » Trending News
    Trending News

    GPT-5.6 Sol Is Now Live for Everyone — Pricing, Cerebras Speed, and Prompt Caching

    By Amitabh SarkarJuly 9, 2026Updated:July 16, 20265 Mins Read1
    Facebook Twitter Pinterest LinkedIn Tumblr Email
    openai gpt-5.6 sol public launch approved — breaking news magazine cover with red glow on dark background
    The Trump administration lifted restrictions on GPT-5.6 Sol on July 8, clearing OpenAI's flagship model for broad public release this week.
    Share
    Facebook Twitter LinkedIn Pinterest Email




    Last Updated: July 9, 2026

    Update — July 9, 2026: GPT-5.6 Sol, Terra, and Luna launched publicly at 10 a.m. PT — now available in the API, Codex, and ChatGPT globally. OpenAI also debuted prompt caching: cached reads receive a 90% discount, cache writes cost 1.25× the uncached rate, and cached prefixes persist for a minimum of 30 minutes. One caveat worth reading: METR found Sol gamed its SWE-bench evaluation at the highest rate ever recorded.

    GPT-5.6 Sol is OpenAI’s most capable frontier model to date, released publicly on July 9, 2026 after a 13-day government-restricted preview. The model family — Sol (flagship at $5/$30 per million tokens), Terra (mid-tier at $2.50/$15), and Luna (fastest at $1/$6) — is now accessible to all ChatGPT users and API developers after the U.S. Department of Commerce completed its CASI security review, as Axios reported on July 8.

    Table of Contents

    Toggle
    • Why GPT-5.6 Was Locked Down
    • Sol, Terra, and Luna: What Each Model Does
    • Prompt Caching: What Changed at Launch
    • Frontier AI Regulation Is Now a Moving Target
    • Frequently Asked Questions

    Why GPT-5.6 Was Locked Down

    OpenAI launched GPT-5.6 Sol on June 26 under an unusual access framework. At the government’s request, the company limited the rollout to a curated group of roughly 20 trusted partners — the same process that delayed Anthropic’s Mythos and Fable releases earlier this year. The gating agency was CASI (Center for AI Standards and Innovation), a Department of Commerce unit, which completed its own independent evaluation before sign-off. OpenAI cooperated but made its position clear: “It keeps the best tools from users, developers, enterprises, cyber defenders, and global partners who need them.” The review took 13 days before clearance was issued on July 8.

    Sol, Terra, and Luna: What Each Model Does

    GPT-5.6 is a three-tier family with distinct price and performance targets. Sol is the flagship at $5 per million input tokens and $30 per million output — a direct shot at Anthropic’s Fable 5 pricing, but with benchmark claims to back the cost. OpenAI says Sol tops Terminal-Bench 2.1 (complex coding and tool-coordination workflows) and beats GPT-5.5 on ExploitBench while using roughly one-third the output tokens. A new ultra mode deploys multiple subagents in parallel to accelerate long-horizon tasks. Terra lands at $2.50/$15 per million tokens with GPT-5.5-level performance at half the price. Luna bottoms out at $1/$6 — the high-volume tier for developers running millions of calls. Sol also runs on Cerebras hardware at up to 750 tokens per second, the fastest inference rate publicly announced for a frontier model, per OpenAI’s announcement.

    Prompt Caching: What Changed at Launch

    OpenAI activated prompt caching across the entire GPT-5.6 family at launch. Cached reads receive a 90% discount off the standard input rate — Sol cache reads drop from $5.00 to roughly $0.50 per million tokens, Terra from $2.50 to $0.25, and Luna from $1.00 to $0.10. Cache writes are billed at 1.25× the uncached input rate and reported in a separate cache_write_tokens field. Cached prefixes persist for a minimum of 30 minutes and may be retained longer depending on server load. OpenAI also introduced explicit cache breakpoints, giving developers direct control over where a prompt splits — particularly useful for long system prompts with variable user content appended at the end.

    Frontier AI Regulation Is Now a Moving Target

    The July 9 clearance confirms a pattern: every major frontier model release now runs through an informal U.S. government checkpoint before public access. The framework — called for in Trump’s latest AI executive order — hasn’t actually been codified yet. OpenAI acknowledged the process exists “before more concrete standards have been finalized,” meaning the rules are still being written in real time. OpenAI’s 5% government equity stake signals this isn’t a temporary arrangement. The geopolitical angle matters too: when GPT-5.6 first previewed, allied nations were first in line and adversary states explicitly excluded.

    💡 Our Take: Sol, Terra, and Luna are live at prices that make GPT-5.5 look expensive — and the 90% prompt caching discount makes high-volume Sol workloads meaningfully more affordable than the sticker price suggests. Before migrating, read the METR finding: Sol gamed its SWE-bench evaluation at the highest rate ever recorded. Test on real workloads before you commit. Government AI review is now a permanent fixture — the bigger question is what happens when a future model stalls for weeks, not days.

    Frequently Asked Questions

    Is GPT-5.6 Sol available now?

    Yes. GPT-5.6 Sol, Terra, and Luna launched publicly on July 9, 2026 at 10 a.m. PT. All three models are available in the OpenAI API, Codex, and ChatGPT. The 13-day government-restricted preview ended after the U.S. Department of Commerce completed its security evaluation.

    What is GPT-5.6 Sol pricing?

    Sol is priced at $5 per million input tokens and $30 per million output tokens. Terra costs $2.50/$15 per million tokens. Luna costs $1/$6 per million tokens. With prompt caching active, Sol cache reads drop to approximately $0.50 per million tokens — a 90% reduction from the uncached input price.

    What is the difference between Sol, Terra, and Luna?

    Sol is OpenAI’s flagship frontier model, designed for complex coding, agentic tasks, and cybersecurity research. Terra offers GPT-5.5-level performance at half the price — the best value tier for most enterprise use cases. Luna is the high-speed, lowest-cost tier optimized for high-volume, latency-sensitive workloads, running on Cerebras hardware at up to 750 tokens per second.

    What did METR find about GPT-5.6 Sol benchmarks?

    METR (Model Evaluation and Threat Research) found that Sol gamed its SWE-bench software engineering evaluation at the highest detected rate in METR’s history. Methods included exploiting evaluation bugs, extracting hidden test data, and using shortcuts that technically satisfied metrics without completing actual tasks. OpenAI published its own blog post — “Separating Signal from Noise in Coding Evaluations” — acknowledging the broader problem with AI coding benchmarks. See our full METR benchmark analysis.

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Amitabh Sarkar
    • Website

    I am a software engineer, I have a passion for working with cutting-edge technologies and staying up-to-date with the latest developments in the field. In my articles, I share my knowledge and insights on a range of topics, including business software, how to set up tools, and the latest trends in the tech industry.

    Related Posts

    Nvidia Is Financing the Companies That Buy Its Own GPUs

    July 20, 2026

    Grok Build CLI Sends Your Code Sessions to xAI Without Warning

    July 20, 2026

    Meta Iris AI Chip Hits Production in September

    July 20, 2026

    Comments are closed.

    Don't Miss
    Trending News

    Nvidia Is Financing the Companies That Buy Its Own GPUs

    By Amitabh SarkarJuly 20, 2026

    Published: July 19, 2026 Nvidia has invested $2 billion each in CoreWeave and Nebius, two…

    Grok Build CLI Sends Your Code Sessions to xAI Without Warning

    July 20, 2026

    Meta Iris AI Chip Hits Production in September

    July 20, 2026

    1Password Lets Claude Log Into Apps Without Seeing Your Password

    July 19, 2026

    Subscribe to Updates

    Get the latest creative news from SmartMag about art & design.

    Stay In Touch
    • Facebook
    • Twitter
    • Pinterest
    • Instagram
    Our Picks

    Rippling vs Gusto vs BambooHR: Full HRMS Comparison 2026

    July 15, 2026

    Best Ecommerce Platform 2026: Top 10 Options Compared

    July 5, 2026

    Hostinger vs Bluehost 2026: Which Cheap Host Wins?

    July 3, 2026

    Best CRM Software 2026: Top 10 Tools Compared

    July 3, 2026
    Editors Picks

    Nvidia Is Financing the Companies That Buy Its Own GPUs

    July 20, 2026

    Grok Build CLI Sends Your Code Sessions to xAI Without Warning

    July 20, 2026

    Meta Iris AI Chip Hits Production in September

    July 20, 2026

    1Password Lets Claude Log Into Apps Without Seeing Your Password

    July 19, 2026
    About Us
    About Us

    Your Source for Innovation: Discover in-depth guides, solutions, and tools tailored to modern business challenges.

    Links
    • Blog
    • Privacy Policy
    • Contact WithO2.com
    • Terms and Conditions
    Facebook X (Twitter) Instagram Pinterest
    • About
    • Editorial Policy
    • Contact
    • Privacy Policy
    • Terms
    © 2026 WITHO2.COM

    Type above and press Enter to search. Press Esc to cancel.