OpenAI announced GPT-5.6, its strongest model family yet, on June 26, then told almost everyone they couldn’t use it. GPT-5.6 Sol set a new state of the art on the Terminal-Bench 2.1 coding benchmark. Its cheaper siblings, Terra and Luna, cut costs in half or better. And access went to “a small group of trusted partners whose participation has been shared with the government.” That’s the GPT-5.6 restricted release in one line: the best model OpenAI has ever shipped, gated to roughly 20 vetted organizations, with no waitlist and no firm public date.
The GPT-5.6 restricted release is OpenAI’s government-requested limited preview of its newest models, Sol, Terra, and Luna, which launched on June 26, 2026 to about 20 approved partner organizations instead of the general public. The day before launch, Axios reported why: the Trump administration’s Office of the National Cyber Director and its Office of Science and Technology Policy asked OpenAI to hold back the rollout while Washington builds an evaluation process for models with serious offensive cyber potential. OpenAI complied. It was the first time the U.S. government preemptively gated an American frontier model before public release.
Users noticed immediately. TechRadar rounded up the reaction within a day, quoting one line that stuck: “the divide has started.” If your product runs on the OpenAI API, this isn’t a story about a delayed toy. It’s a story about your supply chain.
Last updated: July 2026
Quick answers
Why is GPT-5.6 restricted?
The Trump administration’s Office of the National Cyber Director and Office of Science and Technology Policy asked OpenAI to limit the launch while the government builds a pre-release evaluation process for models with advanced cyber capabilities. GPT-5.6 crossed the “High” cybersecurity risk tier in OpenAI’s own Preparedness Framework, so access starts with vetted partners.
When will GPT-5.6 be available to everyone?
OpenAI says GPT-5.6 Sol, Terra, and Luna will reach general availability “in the coming weeks” from the June 26, 2026 preview launch. Sam Altman told employees he hopes broader access follows within a couple of weeks, which puts the realistic window at mid-to-late July 2026. No exact date is confirmed.
Who are OpenAI’s trusted partners for GPT-5.6?
Roughly 20 organizations whose participation was vetted with the U.S. government. OpenAI hasn’t published the list. Altman told staff the government is approving access “customer by customer” during the preview, and partners use the models through the API and Codex rather than ChatGPT.
What is the GPT-5.6 trusted partners restriction?
The restriction is a government-shaped preview period. Instead of shipping GPT-5.6 to ChatGPT subscribers and API developers on day one, OpenAI limited the launch to a short list of organizations it vetted with the U.S. government, roughly 20 in total, who reach the models through the API and Codex. Everyone else waits.
The family itself is a real generational step, not a point release. In OpenAI’s new naming system, the number marks the generation while Sol, Terra, and Luna mark durable capability tiers. Sol is the flagship. Terra matches GPT-5.5’s performance at half the cost, per OpenAI’s launch announcement, and Luna is the fast, low-cost tier. Sol also introduces a “max” reasoning effort and an “ultra” mode that farms complex work out to subagents.
Pricing was published on day one, which tells you OpenAI expects broad availability soon. GPT-5.6 is priced per million tokens at $5 input and $30 output for Sol, $2.50 and $15 for Terra, and $1 and $6 for Luna, per OpenAI’s announcement. Prompt caching got more predictable too: cache writes bill at 1.25x the uncached input rate, cache reads keep a 90% discount, and cached prompts now hold for a 30-minute minimum. OpenAI is also putting Sol on Cerebras hardware in July at up to 750 tokens per second for select customers.
| Model | Input / 1M tokens | Output / 1M tokens | Positioning | Best for |
|---|---|---|---|---|
| GPT-5.6 Sol | $5.00 | $30.00 | Flagship, “max” reasoning and “ultra” subagent mode | Hard agentic coding, security research, long-horizon tasks |
| GPT-5.6 Terra | $2.50 | $15.00 | GPT-5.5-level performance at half the price | Everyday production workloads |
| GPT-5.6 Luna | $1.00 | $6.00 | Fastest and cheapest of the family | High-volume, latency-sensitive features |
Why is GPT-5.6 restricted?
Because of what it can do to computer systems. On June 25, Axios reported that the Office of the National Cyber Director, led by Sean Cairncross, and the Office of Science and Technology Policy, led by Michael Kratsios, asked OpenAI to limit the release. The concern: a model this capable could help automate vulnerability discovery, malware development, and cyberattacks faster than defenders can respond.
The numbers behind that concern are unusually concrete. Reporting based on benchmark data puts GPT-5.6 Sol at 96.7% on OpenAI’s internal capture-the-flag evaluation, a test of autonomous hacking ability, and at 88.8% on Terminal-Bench 2.1. All three GPT-5.6 models crossed the “High” risk classification in OpenAI’s Preparedness Framework, and that assessment, shared with the administration, is what triggered the request. This wasn’t a random regulatory swipe. It was a capability threshold being crossed.
OpenAI’s own disclosures support the dual-use picture. In its launch post, the company says GPT-5.6 Sol identified bugs and exploitation primitives in Chromium and Firefox evaluations but didn’t autonomously produce a full-chain exploit, keeping it under the “Cyber Critical” line. OpenAI also disclosed it spent over 700,000 A100-equivalent GPU hours on automated red teaming to hunt for universal jailbreaks before launch, a figure that gives you a sense of how seriously the company took the risk surface.
OpenAI still made its objection plain: “We don’t believe this kind of government access process should become the long-term default.” The company complied anyway, betting that cooperation is the fastest route to general availability. It’s a pragmatic move for a company already managing a delicate 2026, but it set a precedent no one in the industry can un-ring.

When will GPT-5.6 be available to everyone?
The honest answer is a window, not a date: mid-to-late July 2026 is the realistic target for GPT-5.6 general availability, based on OpenAI’s “coming weeks” language from June 26 and Altman’s internal comment that he hopes broader access lands within a couple of weeks. Nobody outside OpenAI and the administration knows more precisely, because the review process deciding it has no published timeline.
The rollout is staged. Trusted partners get API and Codex access now. OpenAI says ChatGPT, Codex, and broader API availability come “soon,” and the Cerebras deployment of Sol arrives in July for select customers. Watch two signals if you’re planning around it: whether OpenAI expands the partner list (Altman says approvals are happening customer by customer), and whether the White House publishes actual criteria for its cyber Executive Order framework. The second matters more. A defined process gives every future launch a predictable clock. An undefined one means “coming weeks” can stretch.
One more wrinkle: don’t assume the GA models behave identically to the preview. OpenAI says it’s using the preview to tune safeguards that sometimes block or slow legitimate work, especially in security-adjacent domains. Expect some of that friction to survive into general availability. OpenAI says it’s also building longer-term options with enterprise customers, including privacy-preserving detection and customer-operated safety controls, a hint at where differentiated access goes next.
Is Anthropic’s Fable 5 restricted too?
Yes, and Anthropic’s version was far more severe. On June 12 at 5:21 p.m. ET, the Commerce Department issued an export-control directive that forced Claude Fable 5 offline entirely, the first forced shutdown of a commercially deployed frontier model. It stayed dark for more than 17 days. On June 26, Commerce Secretary Howard Lutnick sent Anthropic a letter partially restoring the higher-security Mythos 5 sibling for roughly 100 U.S. organizations that run critical infrastructure, and reporting from June 27 to 29 says full restoration is close, pending Pentagon and NSA sign-off.
Between June 12 and June 25, 2026, the U.S. government restricted frontier model releases from both Anthropic and OpenAI, while Google’s Gemini 3.5 Pro cleared a July launch untouched. That contrast is the tell. Gemini 3.1 Pro scored 70.7% on Terminal-Bench 2.1, more than 18 points below Sol’s 88.8%, and Google’s models haven’t crossed the informal cyber threshold the government appears to be watching. Gemini 3.5 Pro, with its expected 2-million-token context window and reported pricing around $15 input and $60 output per million tokens, is now the only major frontier launch this summer without a government gate.
The difference in how the two restricted labs got treated matters too. Anthropic got a post-launch shutdown and public fight. OpenAI, watching that play out, negotiated a controlled preview in advance. Same threshold, two very different outcomes, and a lesson founders should note: in a rules-free process, the negotiated path beat the adversarial one.
What does the GPT-5.6 restricted release mean for startups?
It means model access is now a dependency risk you have to price, the same way you price cloud outages or App Store review. Thousands of startups build products whose core feature is the latest model’s reasoning. When a frontier release gets gated, companies not on the approved list lose access and predictability overnight, while the roughly 20 partners, mostly large vetted organizations, get weeks of head start. That asymmetry is why the “divide has started” line resonated with developers.
The deeper problem is that the process has no rules yet. Dean Ball, a former White House AI adviser now joining OpenAI, told TechCrunch the arrangement amounts to a “de facto involuntary licensing regime”: no statute from Congress, no published standards, no appeals process, and no defined review length. For a founder, that translates to a roadmap input you can’t forecast. You don’t know which future model gets gated, for how long, or on what grounds.
There’s a competitive dimension as well. If your rival sits on the partner list, or simply builds on an agentic stack that swaps models freely, they compound an advantage while you wait. And if you’ve bet your pitch deck on capabilities only the newest model delivers, an access gap lands directly on your fundraising narrative. Platform dependency has burned builders before, as Meta’s AI turmoil showed founders in a different form. The new version is sharper: the platform itself may be willing, and the government still says wait.
How founders can hedge model-access risk
The correct response to the GPT-5.6 restriction isn’t panic, it’s architecture: multi-model routing, a tested fallback, and no roadmap promises tied to models you can’t call today. Five concrete moves, in priority order.
Put a routing layer between your product and any single model
Route every model call through one internal interface, whether that’s a thin in-house wrapper or an off-the-shelf router like LiteLLM or OpenRouter. The goal is that swapping GPT-5.5 for Terra, or Terra for a Gemini or Claude model, is a config change instead of a rewrite. Teams that had this during the Fable 5 shutdown switched providers in hours. Teams that didn’t spent two weeks refactoring mid-outage.
Never make an unreleased model your core differentiator
Build and demo on models that are generally available today, and treat launch-day access to anything newer as a bonus, not a plan. GPT-5.6 proved that even a shipped model can be inaccessible. If your product only works when Sol-level reasoning arrives, your launch date now belongs to a review process nobody can see into.
Do the token math on redundancy before you need it
Multi-provider redundancy costs less than it used to, and the GPT-5.6 pricing sheet makes the math easy: Terra at $2.50 input and $15 output undercuts GPT-5.5 by half, and 90% cached-read discounts reward stable prompt architecture. Companies are already spending real budget lines on AI tokens, so model your costs across two providers now and you’ll know your worst-case switching bill before a gate forces the question.
Keep an open-weight fallback warm
An open-weight model you host, even at 70% of frontier quality, keeps your SLA alive during a shutdown that takes a hosted provider dark. Fable 5 customers learned this the hard way in June: the model vanished mid-contract, refunds followed, and products built solely on it simply stopped working. A degraded answer beats a 503 page.
Read staged rollouts as strategic signals
The preview shipping through the API and Codex first tells you OpenAI is prioritizing agentic and coding workloads for its most capable tier. That’s useful competitive intelligence for anyone building developer tools or advising companies on AI adoption. Track the partner-list expansion cadence and the framework criteria the White House publishes. The rollout pattern of GPT-5.6 is the template for the next gated launch.

Is government gatekeeping the new normal for frontier AI?
For the most capable models, yes, at least for now. The June 2 executive order created a voluntary pre-release review lane, and within a month it produced a forced shutdown at Anthropic and a gated launch at OpenAI. As of late June, the White House was still working with labs to define what criteria trigger a restriction, which agencies hold authority, and how long a review can run. Congress hasn’t voted on any of it. The threshold is real but unwritten, which is the most uncomfortable kind of rule to build a business under.
Founders have seen this movie in other corners of 2026: federal requirements landing on fast-moving industries before the compliance playbook exists, the way new KYC rules hit stablecoin issuers this spring. The AI version cuts deeper because the input being regulated, frontier model capability, is the raw material of thousands of product roadmaps. Model availability is now a policy variable, not just a commercial one.
The practical read: assume every future frontier launch from a U.S. lab carries a nonzero chance of a gate, and architect so that a two-to-six-week access gap is an inconvenience instead of an existential event. The models will ship. Sol, Terra, and Luna will reach general availability, probably before August. But GPT-5.6 already made the lasting point: the launch calendar for frontier AI now runs through Washington as much as through San Francisco.



