GPT-5.6 Release Date, Features, Pricing & Access Guide
Guides & Insights

GPT-5.6 Release Date, Features, Pricing & Access Guide

Author

Rowan Sterling

Date Published

Latest models · 20View all models
Benchmarks: Artificial Analysis · updated daily
Back to all posts

GPT-5.6 Sol just got an Ultrafast speed tier — up to 14X faster than Standard processing, previewed on the Open​AI API this week — while GPT-5.6 Terra and GPT-5.6 Luna continue the family's rollout into ChatGPT. Two months after the API launch, the GPT​-5.6 generation is both broadly available and still gaining capabilities: the same flagship that pushes past GPT-5.5 on reasoning can now be previewed at up to 750 output tokens per second. The same week, Open​AI also expanded its Daybreak cybersecurity program into Daybreak Blue for defenders and Daybreak Red, powered by a new GPT​-5.6-Cyber model built on Sol for authorized vulnerability research.

Open​AI launched the GPT​-5.6 model family on July 9, 2026, and on August 6 it pushed the flagship into ChatGPT with a reliability-focused update and expanded free access. This is not just another single-model upgrade. It is a tiered family designed to give developers clearer choices across intelligence, speed, and cost — and it is now also the default experience for free ChatGPT users.

The catch is that “GPT​-5.6” means different things in different places.

The API has been live since July 9 with public pricing for all three tiers. The ChatGPT rollout is phased: the updated Sol reached Plus and Pro users on August 6, GPT-5.6 Luna becomes the default for Free and Go accounts this week, and text-only chats on those tiers become unlimited the week of August 10. The Sol behind ChatGPT’s Work and Codex surfaces was not changed by the update. The Ultrafast preview is separate again: a new speed tier in the API, announced August 13, initially limited to a small group of customers.

So GPT​-5.6 is official, shipping, and broadly reachable — and its flagship is now the fastest frontier model on paper. For developers the interesting part is still the routing decision: Sol for the hardest reasoning, Terra for balanced production work, Luna for high volume — and, when the preview opens up, an Ultrafast tier for latency-critical Sol workloads — with prices that changed for two of the three tiers in late July.

This guide explains what GPT​-5.6 is, what is officially confirmed, how Sol, Terra, and Luna differ, how pricing works, what the Ultrafast preview means, and why GPT​-5.6 makes model routing more important.

Quick answer

Is GPT​-5.6 released? Yes. The API launched July 9, 2026, and ChatGPT access is rolling out — updated Sol for Plus/Pro since August 6, Luna as the free/Go default this week.

Can I use GPT​-5.6 in ChatGPT? Yes. GPT-5.6 Sol is now the updated default in the Plus/Pro Chat experience (with a new thought slider), and GPT-5.6 Luna is becoming the default model for Free and Go accounts this week.

Can I use GPT​-5.6 through the API? Yes. Sol, Terra, and Luna are all live on the public API with published pricing, and on OrcaRouter at 0% markup.

What is GPT-5.6 Sol Ultrafast mode? It is a previewed speed tier on the API that runs GPT-5.6 Sol up to 14X faster than Standard processing, at up to 750 output tokens per second. It is initially limited to a small group of customers, and Open​AI has not published pricing for it.

Is there a waitlist? For base access, no — the API is public and ChatGPT access is rolling out across all tiers, so there is nothing to join. Ultrafast mode is different: it is a limited preview, initially open to a small group of customers, with a sign-up for notification as capacity expands.

When is GPT​-5.6 generally available? It already is, in stages: the API went live July 9, 2026, updated Sol reached Plus/Pro ChatGPT on August 6, Luna becomes the free/Go default this week, and unlimited free text chats arrive the week of August 10.

Should you wait for GPT​-5.6? Not if you want to use it — it is live. The real question is which tier to route each task to, and how much reasoning to allow, because that is now what determines your per-task cost.

What is GPT​-5.6?

GPT​-5.6 is Open​AI’s newest model family and the successor generation to GPT-5.5.

The key word is family.

Instead of launching one model for every use case, Open​AI is introducing three tiers:

- GPT-5.6 Sol is the flagship model for the hardest work: deep reasoning, agentic coding, research, professional knowledge work, and complex tool-heavy tasks.

- GPT-5.6 Terra is the balanced model for everyday work. It is designed to offer strong capability at lower cost, making it more suitable for production apps that cannot send every request to the flagship model.

- GPT-5.6 Luna is the fastest and most cost-efficient tier. It is meant for high-volume tasks, lightweight automation, summaries, classification, routing, and other workloads where speed and cost matter more than maximum reasoning depth.

This changes the main question developers need to ask.

The question is no longer: Should I use GPT​-5.6?

The better question is: Which GPT​-5.6 model should handle this task?

Not every request needs Sol. Some tasks need maximum reasoning. Some need lower latency. Some need the cheapest reliable model. GPT​-5.6 makes that tradeoff explicit.

What’s officially confirmed?

GPT​-5.6 should not be treated as a rumor roundup. Open​AI has officially confirmed the model family and preview access.

The confirmed points are:

- GPT​-5.6 is a model family with Sol, Terra, and Luna.

- It launched on the API on July 9, 2026, with all three tiers publicly priced.

- ChatGPT access is rolling out: updated Sol for Plus/Pro (August 6), Luna as the free/Go default (this week), unlimited free text chats (week of August 10).

- Open​AI reports the updated Sol makes ~68% fewer factual errors than GPT-5.5 Instant on internal financial/medical/legal evals (vendor-reported; ~62% for Luna).

- Plus and Pro users get a new thought slider to control how much reasoning each response spends.

- API pricing is public — and Terra and Luna were cut in late July 2026 (Sol unchanged).

- GPT​-5.6 introduces more predictable prompt caching.

- GPT-5.6 Sol includes a new max reasoning effort.

- GPT​-5.6 introduces an ultra mode using subagents for complex work.

- Open​AI previewed Ultrafast mode for GPT-5.6 Sol on the API on August 13, 2026 — up to 14X faster than Standard processing at up to 750 output tokens per second, initially for a small group of customers. The speed figures are vendor-stated and not yet independently reproduced.

- Open​AI expanded Daybreak, its vetted cybersecurity program, into two tiers on August 10, 2026: Daybreak Blue (GPT-5.6 Sol with cyber guardrails removed, for defensive work) and Daybreak Red (GPT​-5.6-Cyber, a Sol-based model purpose-trained for authorized vulnerability research). Access is restricted to vetted customers and security partners, not the public API.

The open questions are different:

- How will the ChatGPT-tuned Sol compare to the API Sol on real tasks?

- How fast will it be in real production workflows? The Ultrafast preview is the first real data point, and it is still invite-only.

- How often will safeguards add friction?

- How will Sol, Terra, and Luna compare on real tasks?

- How much reasoning is enough, and how do you meter the new thought slider?

- How will the Ultrafast tier be priced, and how fast can Cerebras supply keep up with demand for it?

That is why the right approach is preparation, not blind migration.

GPT​-5.6 release date and access

The simplest answer: GPT​-5.6 launched on the API on July 9, 2026, and ChatGPT access is rolling out in phases through August.

API access is now open to any developer with published pricing. What still varies is surface: the Sol tuned for ChatGPT’s Chat experience is not the same as the API model — Open​AI says the version powering Work and Codex is unchanged — so behavior can differ between the consumer app and the API.

For ChatGPT users, the June picture is reversed. Plus and Pro users get the updated Sol with a new thought slider on web, mobile, and desktop that trades response depth against latency. Free and Go users get Luna as the default this week, plus a Think button for harder questions, with unlimited text-only chats arriving the week of August 10.

The newest part of the access story is the Ultrafast preview. Announced August 13, 2026, it is a new speed tier in the API for GPT-5.6 Sol — up to 14X faster than Standard processing (vendor-stated) — and it is starting with a small group of customers rather than an open launch. Open​AI says it is testing the tier across coding, commerce, financial research, customer support, and interactive applications to learn where the speed creates real value before expanding access. Other customers can sign up to be notified when capacity grows.

For developers, the safe phrasing is now: GPT​-5.6 is a shipped, priced API family. The remaining uncertainty is behavioral — how much each tier thinks, how verbose it is, and what that costs per completed task — not whether you can get access. The Ultrafast tier is the one place where access is still the question.

That distinction matters. Do not plan a launch around a GPT​-5.6 access date; plan around tier behavior and per-task cost. Build your evaluation set, meter reasoning effort, and keep a fallback route.

GPT-5.6 Sol, Terra, and Luna explained

GPT-5.6 Sol

Sol is the model to test when quality matters more than cost.

Use it for:

- complex software engineering

- agentic coding workflows

- long-horizon tool use

- difficult debugging

- research analysis

- high-stakes enterprise reasoning

- cybersecurity defense and vulnerability research

Sol is the flagship tier, but that does not mean it should handle every request.

On August 6, 2026, Open​AI also retuned the Sol running in ChatGPT’s Chat experience for everyday conversation — more reliable with facts, more direct answers, and a consistent tone across Instant and Thinking. Open​AI’s internal evaluations, not an independent lab, put factual errors about 68% lower than GPT-5.5 Instant on financial, medical, and legal prompts. The ChatGPT tuning is separate from the API model; the Sol behind Work and Codex is unchanged.

Sol is also the tier getting the Ultrafast preview. On August 13, 2026, Open​AI previewed an Ultrafast mode for Sol on the API that runs the same model up to 14X faster than Standard processing (vendor-stated), aimed at latency-critical work. It is not yet broadly available — the preview starts with a small group of customers — but it is the first sign that Sol’s quality does not have to come with Standard-tier latency.

Sol is also the model behind Open​AI's Daybreak program: Daybreak Blue serves GPT-5.6 Sol with cyber guardrails removed to vetted defenders, and GPT​-5.6-Cyber is purpose-trained on Sol for authorized vulnerability research. Both tiers are restricted-access and sit outside the public API.

GPT-5.6 Terra

Terra is the balanced option.

Use it for:

- everyday AI assistants

- business writing and analysis

- internal knowledge tools

- customer support workflows

- coding help that does not need maximum reasoning

- document processing

- medium-complexity automation

Terra may become the most practical GPT​-5.6 tier for many production apps because it balances capability and price.

GPT-5.6 Luna

Luna is the fast and affordable tier.

Use it for:

- high-volume requests

- classification

- routing

- summarization

- simple transformations

- lightweight agent steps

- fallback responses

- draft generation

Luna is useful when the task is simple enough that paying for Sol would be wasteful.

Luna is also the consumer default now: Open​AI made GPT-5.6 Luna the model behind Free and Go ChatGPT accounts in the week of August 3, 2026, replacing GPT-5.5 Instant. A new Think button lets those users ask for higher reasoning, and text-only chats become unlimited the week of August 10.

GPT​-5.6 pricing and caching

GPT​-5.6 pricing is already public.

GPT-5.6 Sol

- Input: $5 per 1M tokens

- Output: $30 per 1M tokens

- Best for flagship reasoning, complex coding, and deep agentic work

GPT-5.6 Terra

- Input: $2 per 1M tokens (cut from $2.50 in late July 2026)

- Output: $12 per 1M tokens (cut from $15)

- Best for balanced everyday work

GPT-5.6 Luna

- Input: $0.20 per 1M tokens (cut from $1)

- Output: $1.20 per 1M tokens (cut from $6)

- Best for fast, affordable, high-volume workloads

Open​AI cut Terra 20% and Luna 80% in late July 2026, citing inference-efficiency gains; Sol stayed at $5/$30. Because OrcaRouter passes provider pricing through at 0% markup, those cuts show up on our model pages the same day they ship.

The Ultrafast preview has no published price yet. It is a limited preview for a small group of customers, and Open​AI has not said how the tier will be billed — so do not assume the $5/$30 Sol rate applies to it. That is worth watching alongside the speed numbers: if a premium speed tier lands at a premium rate, the cost-per-token math for latency-critical workloads changes.

The pricing makes one thing clear: GPT​-5.6 is not meant to be used as one default model for every task.

A sensible routing strategy might look like this:

- use Luna for simple, high-volume tasks;

- use Terra for everyday reasoning and business workflows;

- use Sol for the hardest coding, research, or agentic tasks;

- cache repeated context aggressively;

- avoid sending every request to Sol by default.

GPT​-5.6 also introduces more predictable prompt caching, including explicit cache breakpoints and a 30-minute minimum cache life. For long-running workflows, coding agents, and enterprise assistants, this can matter a lot.

The practical takeaway: Measure cost per completed task, not just cost per token.

A more expensive model may be cheaper if it finishes the job in fewer attempts. A cheaper model may be better when the task is simple and high-volume.

What’s new in GPT​-5.6?

GPT​-5.6 should not be described mainly through rumored context-window numbers or leaked internal logs.

The important changes are the ones Open​AI has confirmed.

A three-tier model family

Sol, Terra, and Luna make model selection more structured. Developers can route tasks based on difficulty, latency, and cost instead of sending everything to one flagship model.

Stronger agentic and coding workflows

GPT-5.6 Sol is positioned around agentic coding, command-line workflows, planning, iteration, and tool coordination.

That matters because modern coding agents do more than write snippets. They inspect repositories, run commands, understand test failures, patch files, verify changes, and repeat until the task is complete.

New max reasoning effort

Sol introduces a new max reasoning effort for harder tasks.

This should not be used everywhere. Higher reasoning can increase cost and latency, so it should be reserved for work where deeper planning is worth it: complex debugging, migrations, research synthesis, security review, or difficult data analysis.

Ultra mode with subagents

GPT​-5.6 introduces an ultra mode that uses subagents for complex work.

This points toward a future where models do not just answer prompts. They coordinate work across multiple steps or specialized agents.

Ultrafast mode: GPT-5.6 Sol at up to 14X speed

On August 13, 2026, Open​AI previewed Ultrafast, a new service tier in the API that runs GPT-5.6 Sol up to 14X faster than Standard processing, generating up to 750 output tokens per second. Open​AI says the tier delivers the same intelligence as GPT-5.6 Sol Standard — the speedup is about latency, not a smaller or weaker model. Those speed figures are vendor-stated: Open​AI announced them alongside its inference partner Cerebras, and no independent lab has reproduced them yet.

The speed comes from a different inference stack, not a different model. Ultrafast runs on Cerebras’ Wafer-Scale Engine architecture, which keeps model weights on-chip (44 GB of SRAM per wafer-sized chip) and removes the memory-bandwidth bottleneck that limits GPU-based inference. For this tier Open​AI is using Cerebras’ hardware rather than its own.

Ultrafast is initially a limited preview. A small group of customers gets access first, and Open​AI says it is testing the tier across coding, commerce, financial research, customer support, and interactive applications to learn where the speed creates value. Other customers can sign up to be notified when capacity expands. There is no published price and no general availability date yet.

Open​AI and Cerebras also published comparison numbers, all vendor-reported and based on Artificial Analysis output-speed data: Ultrafast claims to be 5X faster than Claude Opus 4.8 in Fast mode and 11X faster than Claude Fable 5; on Humanity’s Last Exam’s 2,500-question set it is said to finish in just over 11 hours versus more than three days for Claude Fable 5, at comparable accuracy; and on GDP-Val (legal briefs, financial models, engineering reports) it claims a 5.6X end-to-end speedup with no loss in quality. Treat those as Open​AI’s and Cerebras’ claims until independent runs confirm them.

The tier also raises two open questions. First, capacity: Cerebras’ wafer-scale supply is the constraint, and whether the preview stays invite-only past late 2026 will signal whether hardware can keep up with demand. Second, pricing structure: a premium speed tier could segment API pricing in ways that matter for cost-sensitive teams.

Stronger safeguards

GPT​-5.6 also comes with stronger safeguards, especially around cybersecurity and biological workflows. Some requests may take longer or return no content while additional safety checks run.

For legitimate defensive or research teams, this does not make GPT​-5.6 unusable. But it does mean sensitive-domain workflows should be tested carefully before production migration.

Daybreak Blue and Red: GPT​-5.6 with guardrails lowered for vetted teams

The flip side of stronger safeguards is that Open​AI now also offers reduced-guardrail access to the same family. On August 10, 2026, it expanded Daybreak — its vetted cybersecurity program — into two tiers. Daybreak Blue gives approved defenders GPT-5.6 Sol with system-level cyber guardrails removed, for vulnerability discovery, secure code review, malware analysis, detection engineering, incident response, and patch validation. Open​AI describes it as the recommended starting point for most defenders.

Daybreak Red goes further. It provides access to GPT​-5.6-Cyber, a model purpose-trained on GPT-5.6 Sol for zero-day discovery and exploit-chain development. Access is heavily restricted and monitored — identity verification, approved-use restrictions, and legal attestations — and Open​AI says individual Daybreak accounts must adopt hardware security keys from September 1, 2026.

The capability figures are vendor-reported. On Open​AI's internal Advanced Cybersecurity Completion Rate, GPT​-5.6-Cyber completed 95% of the sensitive tasks it was asked to do, versus 57.3% for its predecessor GPT​-5.5-Cyber — while Sol through Daybreak Blue completed 2.0% (1.5% with standard safeguards). Open​AI frames the change as a refusal-rate shift more than an intelligence jump: the underlying model answers instead of declining. Open​AI also says the model helped surface two previously unknown vulnerabilities in Chrome's V8 engine, disclosed to Google and fixed as CVE-2026-15903.

Daybreak is not a public API product. Access runs through Open​AI's program and a partner network of security vendors, so most teams will meet it indirectly — in better defensive tooling, and in the signal that Open​AI now ships a model whose job is to attack systems on behalf of defenders.

Sol in ChatGPT: a reliability tune-up and a thought slider

The August 6, 2026 announcement was about the consumer experience. Open​AI tuned the Sol running in ChatGPT’s Chat experience to be more reliable with facts — dates, numbers, sources, rules, and assumptions — and more direct in everyday replies, while bringing the Instant and Thinking experiences closer together. The numbers behind “more reliable” are Open​AI’s own internal evaluations, not an independent lab: on financial, medical, and legal prompts, responses with factual errors were about 68% less common with Sol (and 62% less with Luna) than with GPT-5.5 Instant.

Plus and Pro users also get a new thought slider on web, mobile, and desktop, choosing how much reasoning each response spends — from quick answers for everyday questions to deeper planning, research, writing, coding, or decisions. Open​AI says the version of Sol behind Work and Codex is unchanged; the tuning applies to the Chat experience.

Free and Go: Luna by default, unlimited text chats

Open​AI made GPT-5.6 Luna the default model for Free and Go ChatGPT accounts this week, replacing GPT-5.5 Instant. Starting the week of August 10, text-only chats on those tiers are unlimited — no rate limits, subject to abuse guardrails — while limits still apply to file uploads, images, image generation, and voice. A new Think button gives Free and Go users a way to ask Luna to reason harder, also without limit for text-only chats. Open​AI also added safety training for users under 18, including boundaries around romantic roleplay and age-restricted content.

GPT​-5.6 vs GPT-5.5

Compared with GPT-5.5, GPT​-5.6 matters most in six areas:

- Model choice: GPT​-5.6 introduces Sol, Terra, and Luna instead of one default model.

- Agentic work: Sol is positioned more strongly for coding agents and tool-heavy workflows.

- Reasoning control: max reasoning effort and ultra mode give developers more options for difficult tasks.

- Cost flexibility: Terra and Luna make it easier to avoid overusing the flagship model.

- Speed options: GPT-5.6 Sol previews an Ultrafast tier at up to 14X Standard speed; GPT-5.5 shipped a single speed class.

- Rollout strategy: GPT​-5.6 arrived in phases — API first, then ChatGPT — so teams still need fallbacks as behavior and pricing shift.

The key point is simple: GPT​-5.6 is not just a better model. It is a reason to stop hard-coding one model into your app.

How to prepare for GPT​-5.6

You can prepare with access you already have: Luna is the free ChatGPT default, Sol is in the Plus/Pro experience, and the API is public. The preparation is about routing and metering, not waiting.

First, build a real evaluation set. Do not test GPT​-5.6 only with toy prompts. Use actual codebase issues, support tickets, customer documents, research tasks, and automation failures.

Second, design for model routing. Luna can handle simple summaries and classification. Terra can handle everyday user-facing workflows. Sol should be reserved for hard reasoning, coding, and research. If latency is a requirement on your roadmap, treat the Ultrafast preview as a reason to design a speed-aware routing layer now rather than hard-coding one tier — the tier that wins on latency today may still be in preview, but the routing decision it implies is already visible.

Third, track cost per completed task. For agentic workflows, one user request can trigger many model calls. The cheapest token price is not always the cheapest completed workflow.

Fourth, prepare for caching. If your app repeatedly sends the same repository, policy document, or knowledge base, structure prompts so stable context can be reused.

Finally, keep a fallback plan. Reasoning models are verbose, and prices move — Terra and Luna both dropped in late July. Production systems should be able to fall back to another GPT​-5.6 tier, to GPT-5.5, or to a different provider when latency, cost, or behavior shifts.

Where OrcaRouter fits

GPT​-5.6 makes model routing more important.

Open​AI is no longer presenting the new generation as a single model. GPT​-5.6 comes as a family: Sol for the hardest reasoning and agentic work, Terra for balanced everyday workloads, and Luna for fast, lower-cost tasks.

That creates a practical question for developers: Which model should handle each request?

Not every task needs Sol. A simple summary may be better served by Luna. A normal business workflow may fit Terra. A difficult coding-agent task may justify Sol.

And even with access open, production apps still need fallback options when latency, rate limits, or per-task cost move — the late-July price cuts are a reminder that the rate card is a moving target, and the Ultrafast preview is a reminder that the serving options are moving too.

This is where OrcaRouter fits.

OrcaRouter gives developers one Open​AI-compatible endpoint for routing across many models, with smart routing, automatic failover, observability, and zero token markup. Instead of rebuilding your integration every time a new model launches, you can keep one API pattern and route tasks based on quality, cost, latency, and availability.

For GPT​-5.6, that means you can build the model layer now rather than waiting for access to settle.

All three standard tiers are on OrcaRouter today: GPT-5.6 Sol at $5/$30, GPT-5.6 Terra at $2/$12, and GPT-5.6 Luna at $0.20/$1.20 — passed through at 0% markup, so the late-July price cuts are live here the day they shipped. The Ultrafast tier is not broadly available yet, so it is not part of that list — when it opens up, it becomes one more serving tier to route across, and the same endpoint is where that choice lives. Daybreak Blue and Red are a separate restricted-access program rather than public API tiers, so they are not routable through OrcaRouter either — the standard tiers are where the routing decision lives.

GPT​-5.6’s access story is still settling, but your routing strategy does not have to wait.

Using GPT​-5.6 through OrcaRouter

GPT​-5.6 is exactly the kind of family developers want to route across.

But the rollout has created a practical problem: access expanded in phases, prices changed for two tiers in late July, the ChatGPT tuning differs from the API model, and now a speed tier is in preview. Teams need a layer that absorbs those changes.

OrcaRouter exposes all three GPT​-5.6 tiers through one Open​AI-compatible endpoint, alongside 200+ other models — so you can test Sol, Terra, and Luna against GPT-5.5, Claude, Gem​ini, and the rest in one place, without re-integration.

This is useful if you want to:

- route each request to the cheapest tier that handles it;

- test Sol, Terra, and Luna side by side with your current stack;

- route simple tasks to cheaper models;

- reserve flagship models for harder work;

- compare GPT​-5.6 with your current production stack;

- keep fallback models ready when behavior or pricing shifts.

GPT​-5.6 is live on OrcaRouter now — Sol, Terra, and Luna — at 0% markup, with automatic failover so a newly shipped or repriced tier never takes down a production path.

Call any GPT​-5.6 tier through one API on OrcaRouter, and route across the rest of the catalog with the same key.

Final verdict

GPT​-5.6 is real, official, and important.

But it is no longer a preview you wait for: the API has been live since July 9, and ChatGPT access is rolling out to every tier this month.

The biggest shift is not just better intelligence. It is the move from one default model to a routed model family: Sol, Terra, and Luna.

For developers, the opportunity is not simply upgrading to GPT​-5.6. The opportunity is building a smarter model layer:

- Luna for fast and cheap tasks;

- Terra for balanced production workloads;

- Sol for the hardest reasoning and coding tasks;

- caching for repeated context;

- fallback models for reliability.

If you are shipping today, you no longer have to wait: Luna is the free ChatGPT default, Sol is in the Plus/Pro experience, and the API is public.

If you are building coding agents, enterprise copilots, or long-running AI workflows, preparation now means routing, reasoning-effort metering, and per-task cost — not access. And the Ultrafast preview is a reminder that the family is still moving: the flagship you benchmark today may have a faster tier — or, with Daybreak, a purpose-trained cyber variant — by the time you ship.

FAQ

Is GPT​-5.6 released?

Yes. The API launched July 9, 2026, and ChatGPT access is rolling out — updated Sol for Plus/Pro since August 6, Luna as the Free/Go default this week.

Can I use GPT​-5.6 in ChatGPT?

Yes. GPT-5.6 Sol is live in the Plus/Pro Chat experience with a new thought slider, and GPT-5.6 Luna is becoming the default for Free and Go accounts this week, with unlimited text-only chats starting the week of August 10.

Can I use GPT​-5.6 through the API?

Yes. All three tiers are on the public API with published pricing — Sol $5/$30, Terra $2/$12, Luna $0.20/$1.20 — and through OrcaRouter at 0% markup.

What is GPT-5.6 Sol Ultrafast mode?

It is a previewed speed tier in the Open​AI API that runs GPT-5.6 Sol up to 14X faster than Standard processing, at up to 750 output tokens per second, with the same intelligence as the standard model (vendor-stated). It is initially limited to a small group of customers, and Open​AI has not published pricing for it.

Is there a public waitlist for GPT​-5.6?

For the base tiers, no — the API is public and ChatGPT access is rolling out across all tiers, so there is nothing to join. The Ultrafast preview is the exception: it is limited to a small group of customers initially, with a sign-up for notification as capacity expands.

When will GPT​-5.6 be generally available?

It already is, in stages. The API went live July 9, 2026; updated Sol reached Plus/Pro ChatGPT on August 6; Luna becomes the Free/Go default this week; unlimited free text chats arrive the week of August 10.

What are GPT-5.6 Sol, Terra, and Luna?

Sol is the flagship model, Terra is the balanced model for everyday work, and Luna is the fast and affordable model.

How much does GPT​-5.6 cost?

GPT-5.6 Sol costs $5 per 1M input tokens and $30 per 1M output tokens. GPT-5.6 Terra costs $2 per 1M input and $12 per 1M output. GPT-5.6 Luna costs $0.20 per 1M input and $1.20 per 1M output. Terra and Luna were cut in late July 2026; Sol was unchanged. On OrcaRouter these prices pass through at 0% markup. The Ultrafast preview does not have a published price yet.

Should I switch from GPT-5.5 to GPT​-5.6?

Not blindly. Keep a stable fallback, but the calculus has changed: Luna’s price makes it the default choice for high-volume text work, and the ChatGPT tuning is worth testing for everyday assistants. Benchmark on your own tasks and route per tier.

Can I access GPT​-5.6 through OrcaRouter?

Yes. GPT-5.6 Sol, Terra, and Luna are all on OrcaRouter through one Open​AI-compatible endpoint at 0% markup, with automatic failover. You can test them against GPT-5.5, Claude, Gem​ini, and other models without changing integrations.

What is Daybreak Blue and Daybreak Red?

They are the two tiers of Open​AI's vetted cybersecurity program, expanded on August 10, 2026. Daybreak Blue gives approved defenders GPT-5.6 Sol with cyber guardrails removed, for defensive work such as incident response, malware analysis, and secure code review. Daybreak Red provides GPT​-5.6-Cyber, a Sol-based model purpose-trained for authorized vulnerability research and exploit work, to separately vetted teams. Neither is a public API product.

Compared in this article4

Detected from this article · Benchmarks: Artificial Analysis · updated daily