OpenAI has officially announced the launch of its next-generation frontier model family, GPT-5.6, marking a seismic shift in artificial intelligence capabilities. This newly introduced suite comprises three distinct, tier-based models inspired by celestial bodies: Sol (the flagship power), Terra (the balanced production worker), and Luna (the hyper-fast, low-cost pipeline).
The premier model of this release, GPT-5.6 Sol, sets a new industry standard—outperforming current market rivals, including Anthropic’s highly acclaimed Claude Mythos 5, across critical operational benchmarks.
A defining breakthrough of this release is the deployment of specialized execution configurations. OpenAI has introduced two advanced reasoning tiers: a max reasoning mode—which grants the model extended compute time to process highly convoluted logic—and a revolutionary ultra mode. The ultra configuration leverages a multi-agent paradigm, deploying a network of subagents to coordinate and accelerate multi-step tasks simultaneously.


On Terminal-Bench 2.1—a benchmark evaluating command-line engineering, complex tool usage, and iterative planning—GPT-5.6 Sol (Ultra) secured a historic, state-of-the-art score of 91.9%. This directly eclipses Claude Mythos 5, which trails at 84.3%.
Shifting the token-efficiency frontier
Beyond raw intelligence, the GPT-5.6 architecture fundamentally changes the economics of large language model deployments by heavily optimizing token consumption.


- Unprecedented Cyber Efficiency: On ExploitBench², GPT-5.6 Sol achieved performance parity with Mythos Preview while consuming only roughly 1/3 of the output tokens.
- Scientific Advancements: On GeneBench v1 (long-horizon genomics and quantitative biology), Sol surpassed the legacy GPT-5.5 flagship while utilizing a substantially smaller token footprint.
Granular pricing structure
Under the newly implemented naming convention, the number (5.6) identifies the overall model generation, while the names denote permanent capability tiers designed to scale independently:
- GPT-5.6 Sol (Flagship): Engineered for advanced vulnerability research, multi-agent pipelines, and deep reasoning. Priced at $5.00 input / $30.00 output per 1M tokens.
- GPT-5.6 Terra (Balanced): Built for efficient, high-volume production. It matches legacy GPT-5.5 capability while being explicitly 2x cheaper ($2.50 input / $15.00 output), comfortably beating Claude Fable 5 in comparative data.
- GPT-5.6 Luna (Fast): Optimized for low-cost, high-velocity data pipelines. Priced at an ultra-low $1.00 input / $6.00 output.
Hardware Acceleration: OpenAI announced an infrastructure partnership with Cerebras, launching GPT-5.6 Sol this July at a blistering speed of up to 750 tokens per second for enterprise clients requiring real-time inference.
Technical enhancements for enterprises and developers
To provide financial predictability for enterprises dealing with massive context windows, OpenAI has overhauled its Prompt Caching mechanics:
- A guaranteed minimum cache lifetime of 30 minutes has been introduced.
- Initial cache writes incur a 1.25x billing premium over standard uncached input rates, but subsequent cache reads enjoy a steep 90% discount, heavily driving down operational costs for continuous codebases.
The most talked-about operational aspect of this release is its limited availability. Due to sovereign security protocols and coordination with the U.S. administration, the GPT-5.6 family is initially launching in a highly restricted preview for a small cohort of trusted partners.
In an unusual move, OpenAI used its official product documentation to publicly critique this sovereign oversight:
“We don’t believe this kind of government access process should become the long-term default. It keeps the best tools from users, developers, enterprises, cyber defenders, and global partners who need them.”
The company framed this phased release as a short-term compromise while working within the framework of the White House’s cyber Executive Order. OpenAI emphasized that GPT-5.6 Sol’s safety stack is its most robust to date. To achieve this, the company dedicated over 700,000 A100-equivalent GPU hours to automated red-teaming to neutralize universal jailbreaks.
During stress tests involving the Chromium and Firefox codebases, Sol proved exceptionally capable at isolating bugs and patch development (defensive work) but failed to autonomously engineer a functional, full-chain exploit—keeping it safely below the lab’s “Cyber Critical” threshold.
Individual consumers and developers won’t have to wait long; OpenAI plans to move the GPT-5.6 suite into General Availability across ChatGPT, Codex, and the public API in the coming weeks.
Source: openai.com
















