Vuncloud Blog
← Back to Dev Notes

What Is GPT-5.6? Sol, Terra & Luna Family, Pricing and Government-Reviewed Launch Explained (2026)

The Sol / Terra / Luna family · how the government-reviewed limited preview unfolded · full API pricing · ChatGPT rollout status · which tier to pick~13 min read

OpenAI GPT-5.6 latest AI model interface, representing the Sol / Terra / Luna family's reasoning and multimodal capabilities

"What actually is GPT-5.6? Why isn't it a single model like every version before it? And why did the internet say the US government held up its release for two weeks?"

If you saw the name "GPT-5.6" and assumed it was just a routine point release of GPT-5.5, the reality is a bigger departure than that. This time OpenAI did three things it had never done before: shipped three capability tiers of models at once instead of one, split the launch into a "limited preview" phase and a "public launch" phase, and for the first time delayed a public release because of a US government cybersecurity review. This article walks through it in order — what it is, why it launched this way, what it costs, and how to choose a tier.

3
New model family: Sol (flagship) / Terra (balanced) / Luna (lightweight)
7.9
Public launch date (July 9, 2026 — today)
$1–$5
API input price range across the three tiers (per 1M tokens)

1. What Is GPT-5.6

GPT-5.6 is the first generation released under OpenAI's new model naming system in 2026: instead of one model mapped to one version number, it's a family made up of three "capability tiers" —

  • Sol: the flagship model, built for top-tier reasoning, coding agents, and long-chain complex tasks in biology and cybersecurity.
  • Terra: the balanced tier, performance roughly on par with GPT-5.5 but at half the price (OpenAI's own phrasing is "2x cheaper"), aimed at large-scale production workloads.
  • Luna: the lightest, fastest, and cheapest tier, built for high-concurrency, low-latency everyday tasks.

OpenAI has explained the logic behind this new naming scheme directly: the number represents the model generation, while Sol / Terra / Luna represent "durable capability tiers" that can each iterate on independent schedules — going forward, Terra or Luna might get their next generation faster than Sol does. It's a departure from the old model of binding a single model to one major version number.

This isn't "GPT-5.5 with a new name"

Terra's positioning is the key detail: OpenAI's own data shows it "performs close to GPT-5.5 at half the price." In other words, what used to require GPT-5.5 can now be done with Terra for half the cost; the real generational leap is Sol's jump in reasoning depth, coding agent ability, and security research capability.

2. Release Timeline: Why the Government Review Delayed It

GPT-5.6's rollout was a rare "staged, government-reviewed" release for OpenAI — a sharp break from its usual "release and it's available" cadence:

Date Event Scope
June 26, 2026 OpenAI publishes the GPT-5.6 Sol limited-preview post, announcing the max/ultra reasoning modes and initial pricing ~20 government-vetted trusted partner organizations, API + Codex only, ChatGPT not included
Late June – early July 2026 OpenAI continues working with the US government and third-party red-teaming organizations to test safety guardrails without widening access Restricted partners only
July 8, 2026 OpenAI's official X account announces Sol / Terra / Luna will launch publicly "this Thursday" and that preview access is "starting to expand globally now" Global preview announcement
July 9, 2026 (today) Official public launch: Sol, Terra, and Luna begin rolling out to ChatGPT, Codex, and API users Global users (rolling out in stages)

The reason behind it: GPT-5.6 showed a clear jump in cybersecurity-related capability (vulnerability discovery, exploit research). Citing national security concerns, the US government required OpenAI to first validate its safety safeguards within a small, controlled, vetted group before deciding whether to widen the release. In its own blog post, OpenAI stated plainly: "we don't think this kind of government access process should become the long-term default", but accepted it as a short-term step toward broader availability, while jointly developing a cybersecurity executive-order framework with the US government for future model releases.

Context worth noting: rival Anthropic's Claude Fable 5 and Claude Mythos 5 previously had their access paused under export-control directives too, and were restored in late June after the US Department of Commerce lifted the restrictions — suggesting this wave of tightened oversight isn't aimed at a single company, but is a new normal the whole frontier-model industry is navigating.

3. Core New Features In-Depth

3.1 Max Reasoning & Ultra Subagent Mode

GPT-5.6 introduces two new reasoning-control concepts, currently focused mainly on Sol:

  • max reasoning effort: a new top tier above the existing reasoning-effort levels, giving Sol the longest thinking time for the hardest reasoning tasks.
  • ultra mode: not "think longer" but "split the work across others" — the model automatically spawns subagents to work on different parts of a complex task in parallel, breaking past the efficiency ceiling of a single agent working sequentially. Third-party evaluations show Sol Ultra scoring 91.9% on Terminal-Bench 2.1, versus 88.8% for base Sol.

Ultra mode isn't cheap — use it selectively

Sol's output price is $30 per 1M tokens, and ultra mode spawns multiple subagents working in parallel, which noticeably increases actual token consumption. It's best reserved for highly parallelizable tasks where correctness matters more than cost; for everyday work, max or the base reasoning effort is usually enough.

3.2 Coding / Biology / Cybersecurity Benchmarks

OpenAI's published evaluations focus on three areas, each emphasizing "doing more with fewer tokens":

  • Coding (Terminal-Bench 2.1): GPT-5.6 Sol sets a new state-of-the-art on this command-line workflow benchmark, which evaluates multi-step planning, iteration, and tool coordination.
  • Biology (GeneBench v1): on long-horizon genomics and quantitative biology analysis tasks, Sol achieves better results than GPT-5.5 while using fewer tokens.
  • Cybersecurity (ExploitBench / ExploitGym): on exploit-related benchmarks, Sol matches Anthropic's Mythos Preview using roughly a third of the output tokens; across Sol, Terra, and Luna, cybersecurity capability improves significantly when reasoning effort is increased (ExploitGym was developed by UC Berkeley researchers in collaboration with OpenAI and other frontier labs).
Multiple monitors displaying code and security-test output, representing GPT-5.6 Sol's cybersecurity and coding benchmark performance
GPT-5.6 Sol's jump in cybersecurity-related benchmark performance is the core reason behind this staged launch

3.3 Stronger Capability, Stronger Safety Guardrails

Alongside the capability jump, OpenAI says it invested in its "most robust safety stack yet":

  • Over 700,000 A100-equivalent GPU hours of automated red-teaming, specifically hunting for cross-context "universal jailbreak" techniques.
  • Layered guardrails: refusal training at the model level, real-time risk classifiers during generation, account-level behavioral review, and differentiated access permissions — stacked together rather than relying on a single line of defense.
  • Per OpenAI's own Preparedness Framework assessment, Sol does not cross the "Cyber Critical" threshold: in tests targeting Chromium and Firefox, the model could find vulnerabilities and exploit primitives, but did not autonomously assemble a complete, working attack chain under test conditions.
  • OpenAI's stated positioning is clear: Sol is better at "helping find and fix vulnerabilities" than at "reliably executing end-to-end attacks" — the goal is for defenders (security researchers, red teams, patch developers) to benefit first.

3.4 Context Window (Not Officially Confirmed)

Worth flagging clearly: neither OpenAI's official launch materials nor its system cards disclose a specific context window length for the GPT-5.6 family. Third-party estimates, extrapolating from the previous generation GPT-5.5 (roughly a 1M-token window), guess 5.6 is similar or close, but that's speculation only — don't hard-code a specific number into production code. Always rely on the model metadata returned by the API.

3.5 New Caching Mechanics

GPT-5.6 also updates the billing rules for prompt caching, making costs more predictable than before:

  • Supports explicit cache breakpoints, with a minimum cache lifetime of 30 minutes.
  • Cache writes: billed at 1.25x the model's uncached input price.
  • Cache reads: continue to get a 90% discount off the uncached input price.

For agent workloads that repeatedly send long system prompts or large context blocks, this makes costs much easier to estimate upfront — as long as the same prefix is reused multiple times within the 30-minute window, the cache-read discount can meaningfully thin out total spend.

4. Full Pricing Table

All three GPT-5.6 tiers are billed per 1 million tokens. This is preview-period pricing, and OpenAI notes it "may be adjusted at general availability":

Model Model ID Input (/1M tokens) Output (/1M tokens) Cache Write (1.25×) Cache Read (-90%)
Sol (flagship) gpt-5.6-sol $5.00 $30.00 $6.25 $0.50
Terra (balanced) gpt-5.6-terra $2.50 $15.00 $3.13 $0.25
Luna (lightweight) gpt-5.6-luna $1.00 $6.00 $1.25 $0.10

Worth noting: Sol's pricing is identical to GPT-5.5's ($5/$30) — meaning the reasoning, coding, and security capability jump comes at no extra cost; Terra's pricing is exactly half of Sol/GPT-5.5, and OpenAI describes its performance as "close to GPT-5.5"; Luna is the cheapest of the three, built for speed and high-volume calls.

Sol also has a "Fast" mode

OpenAI has announced it will run GPT-5.6 Sol on Cerebras hardware starting in July, reaching speeds of up to 750 tokens per second, built for scenarios that demand extreme response speed. It's initially limited to select customers, with capacity expanding over time.

5. ChatGPT / API / Codex Rollout Status

As of today (July 9, 2026), here's the actual status across each channel:

Channel Current Status Notes
API ✅ Public launch underway, expanding gradually Preview was limited to ~20 vetted partners; opening to a broader developer base starting today
Codex ✅ Public launch underway, expanding gradually Rolling out on the same schedule as the API
ChatGPT (web / app) 🔄 Rolling out, not available to everyone at once Not available at all during the preview; opening to subscribers in batches starting today — which plan and region come first has not been officially detailed
Whether ChatGPT Pro splits into Sol / Terra / Luna tiers ❓ Not officially confirmed A third-party benchmark site previously spotted what looked like clues of a three-tier Pro configuration in leaked test material, but OpenAI has not commented — treat this as unconfirmed until an official announcement

If you open ChatGPT and don't see Sol / Terra / Luna options yet, that's expected — this is a "gradual expansion" rollout, not an instant switch-on for everyone. Keep an eye on your account's model picker or OpenAI's official announcements for the latest status.

GPT-Live voice models launched alongside it

On July 8, OpenAI also released its next-generation voice models, GPT-Live-1 and GPT-Live-1 mini, which support simultaneous "listen and speak" interaction that OpenAI describes as feeling "more like a real conversation." This is a separate product line from the GPT-5.6 text/reasoning family, launched around the same time.

6. Key Differences vs GPT-5.5

Dimension GPT-5.5 GPT-5.6 Practical Impact
Product form Single model Three-tier family (Sol / Terra / Luna) Choose by intelligence / speed / cost instead of one-size-fits-all
Flagship price (input/output, /1M tokens) $5 / $30 Sol: $5 / $30 (unchanged) Same price for the flagship, capability gains come free
Balanced tier option No equivalent tier Terra: $2.50 / $15, performance close to GPT-5.5 What used to require flagship pricing now costs half as much for near-equivalent results
Reasoning control Standard reasoning-effort levels New max and ultra (subagent orchestration) A higher reasoning ceiling for the hardest tasks
Caching Implicit caching Explicit cache breakpoints, writes at 1.25×, reads at -90% More predictable cost for long-context/long-session workloads
Release approach Standard direct launch Limited preview (government-reviewed) → staged public launch Sensitive-capability scenarios take longer to reach full access
Cybersecurity capability Baseline level Significantly improved, paired with strengthened safety guardrails Big win for security research/red teams — also the reason for the review

7. Should You Use Sol, Terra, or Luna?

Your Scenario Recommended Tier Why
Complex coding agents, long-chain reasoning, security research Sol (enable ultra when needed) The only tier that unlocks max/ultra reasoning and top benchmark performance
Large-scale production, optimizing for value Terra Officially positioned as "close to GPT-5.5 performance, half the price" — the best fit for replacing existing GPT-5.5 production traffic
High-concurrency, low-latency everyday tasks (classification, summarization, support) Luna The lowest unit price in the family, ideal for scale to keep costs thin
Not sure which one to use, want to try first Start with Terra Good enough performance at a mid-range price — the most balanced default choice
Need extreme response speed (e.g. real-time interaction) Sol (Cerebras Fast mode, select customers) or Luna Fast mode is built for speed but access is limited; Luna is the widely available high-speed option today

8. Developer Integration Notes

8.1 Three Separate Model IDs

You now have to explicitly specify a tier when calling the API — it's no longer a single model name: gpt-5.6-sol / gpt-5.6-terra / gpt-5.6-luna. It's worth building routing logic by task complexity — default simple tasks to Luna, and only escalate to Terra/Sol for complex or high-value work, instead of sending every request to Sol.

8.2 Preview-Period Access Limits May Still Apply

Even though the public launch happened on July 9, some capabilities (like Sol's Cerebras Fast mode or the full functionality of ultra subagent mode) may still be rolling out in stages. Before wiring anything into production, verify in a test account that your target model ID and reasoning-effort parameters are actually enabled for you.

8.3 Don't Hard-Code the Context Window Length

As noted above, OpenAI hasn't published a specific number. Always read the actual context and output limits from the model metadata returned by the API, rather than assuming a fixed value in your capacity-planning code.

8.4 Take Advantage of the New Caching Rules

If your agent repeatedly reuses the same system prompt or long context block, set up explicit cache breakpoints and schedule frequently reused requests within the 30-minute window to benefit meaningfully from the 90% read-side discount.

Pairing Cloud Mac with GPT-5.6 agents

  • Sol's ultra subagent mode is well suited to breaking work into multiple parallel subtasks — if any of those touch the macOS ecosystem (Xcode builds, TestFlight packaging), you'll need a stable execution node to run them on.
  • During the transition from preview to full availability, it's worth running production workflows on Terra first, then evaluating whether Sol is worth the upgrade.
  • Vuncloud's dedicated Mac mini M4, available on monthly rental, is a good fit for running the macOS-related subtasks of a GPT-5.6 agent in a stable environment over long stretches of time.

FAQ

What is GPT-5.6?

Not a single model — it's the Sol (flagship) / Terra (balanced) / Luna (lightweight) three-tier family, the first generation released under OpenAI's new model naming system.

When was GPT-5.6 actually released?

Limited preview on June 26 (API/Codex only, restricted to ~20 government-vetted partners, ChatGPT excluded); official public launch on July 9 (today), with ChatGPT, Codex, and API access opening gradually.

Why did it need a government review?

Sol showed a significant jump in cybersecurity-related capability (vulnerability discovery and exploit research). The US government, citing national security, required safeguards to be validated within a controlled group first; OpenAI accepted this as a short-term transitional step.

What does each tier cost?

Per 1M tokens: Sol $5/$30, Terra $2.50/$15, Luna $1/$6. Sol is priced the same as GPT-5.5; Terra is half of Sol.

Is it available in ChatGPT yet?

Not at all during the preview; rolling out in batches starting July 9. Which plans and regions come first hasn't been officially published — check your account's model menu.

Conclusion

What is GPT-5.6? It's OpenAI's first answer after retiring the "one model per version number" model: the Sol / Terra / Luna family, letting users, for the first time within a single generation, pick precisely along intelligence, speed, and cost.

Why was the launch so unusual? Because Sol's jump in cybersecurity capability triggered a US government review requirement, and OpenAI chose a "limited validation first, public launch second" transition path — this is unlikely to be the last time a frontier model runs into this kind of situation.

How should you choose? Most production use cases can swap straight to Terra in place of GPT-5.5 — close performance, half the price; reach for Sol when you genuinely need top-tier reasoning and agent orchestration; hand high-concurrency lightweight tasks to Luna. Starting today, all three tiers are gradually opening up to everyone — if you don't see them in ChatGPT or your API account yet, give OpenAI a little more time.

Need a stable macOS node for GPT-5.6 Agent tasks?

Whether you're running Sol's ultra subagent orchestration or Terra's day-to-day production workloads, Xcode builds and TestFlight packaging need an execution environment that doesn't drop offline. Vuncloud's dedicated Mac mini M4 is available on monthly rental, ready to use out of the box.

View Cloud Mac Plans · LLM Pricing Guide

This article is based on OpenAI's official blog post and public reporting (9to5Mac, CNBC, The Economic Times, VentureBeat, and others). Preview-period pricing and access scope may change at general availability — refer to OpenAI's latest official announcements. Last updated: July 9, 2026.

Dev Notes · AI Models

GPT-5.6 Sol/Terra/Luna · Cloud Mac Execution Node

Choose the right tier across the three-model family · Dedicated Mac mini M4 · Stable execution for agent tasks

View Cloud Mac Plans
Limited Offer View Plans