A generated hero card headed 'One tier, three models' with three chips reading 'GPT-6.1 Sol Ultrafast: available, $12.00 / $60.00', 'GPT-6 Astra Ultrafast: available, $60.00 / $300.00' and 'GPT-5.6 Sol Ultrafast: preview only, no published rate', over a footer reading 'As read from OpenAI's documentation on 9 October 2026'.
Guides & Insights

GPT-6.1 Sol Ultrafast vs GPT-5.6 Sol Ultrafast: One Tier, Three Models, One of Them Still Waiting

Author

Alistair Wren

Date Published

Latest models · 20View all models →
Benchmarks: Artificial Analysis · updated daily
Back to all posts

GPT-6.1 Sol Ultrafast and GPT-5.6 Sol Ultrafast are the same service tier applied to two different Sol generations, and after October 8, 2026 they are in opposite states. GPT-6.1 Sol went from "coming soon" to a published rate card — $12.00 per million input tokens, $60.00 per million output tokens, six times Standard — available to every API user, and to Codex and ChatGPT Work users on the Pro $500 plan. GPT-5.6 Sol is still where Ultrafast began: an announced tier, in limited preview, with no row on the pricing page at all. Same name, same mechanism, and only one of them is something you can buy.

That is the short version. The longer version matters because the tier's headline speed numbers are attached to different models, measured in different surfaces, and in one case three months stale.

What "Ultrafast" actually is on each model

On neither model is Ultrafast a separate checkpoint. You send the same model id and flip a service tier — model: "gpt-6.1-sol" with service_tier: "ultrafast" in the Responses API — and O​penAI schedules your request on the low-latency path. The weights, the answers and the context window are unchanged. On GPT-6.1 Sol that window is 1,050,000 tokens with a 922,000-token maximum input and 128,000 maximum output; GPT-5.6 Sol carries the identical window and output ceiling, and differs mainly in knowledge cutoff (February 16, 2026 against April 30, 2026) and in the reasoning ladder, where GPT-5.6 Sol still accepts none and GPT-6.1 Sol does not.

So the comparison is not "which model is faster at thinking". It is "which generation of the same tier is open to you", and the answer changed four days ago.

A screenshot of OpenAI's Ultrafast mode documentation page, showing the sentence 'Ultrafast mode is the fastest service tier in the OpenAI API', the availability sentence naming broad availability for GPT-6 Astra and GPT-6.1 Sol with preview access for GPT-5.6 Sol, the WebSockets recommendation warning that network overhead can reduce the latency gains without a persistent connection, the note that Ultrafast has separate rate limits from Standard and Fast modes, the line that GPT-6.1 Sol supports US and EU data residency, and a code sample setting service_tier to 'ultrafast' with model 'gpt-6.1-sol'.

The published multiples, and the one number that stayed at 14x

O​penAI has published two speed claims about this tier, on two models, and neither is a general statement about "Ultrafast".

• GPT-6 Astra Ultrafast — "generates tokens up to 8x faster than GPT-6 Astra in Standard mode in Codex". This is the tier's current headline, repeated on the ChatGPT-side speed documentation.

• GPT-5.6 Sol Ultrafast — announced August 13, 2026 as running "up to 14x faster than Standard processing", in limited preview to select customers.

Notice what is missing. There is no published speed multiple for GPT-6.1 Sol Ultrafast. Its documentation page describes the tier, gives its rate limits and data-residency posture, and points at the pricing table — but the only model-specific multiple in the whole docs set belongs to Astra. The honest position is that "up to 8x" is Astra's measurement, on the same serving stack, and it is the number the rollout is being sold on without having been reproduced for Sol by anyone, inside O​penAI or out.

The 14x number deserves more suspicion than it usually gets, because of how old it is. Announced in mid-August for a tier that has still not reached general availability, it has sat unrevised for roughly two months while the tier itself was rebuilt around two other models. In that time the obvious thing to expect — that 5.6 Sol's Ultrafast would graduate from preview and get priced — never happened. Instead GPT-6 Astra got it on September 29 and GPT-6.1 Sol got it on October 8. A number that only exists in an announcement, attached to a configuration nobody can buy, is marketing copy until it is on a rate card.

Side by side: what each one costs and who can reach it

• Standard rate — GPT-6.1 Sol $2.00 input / $0.10 cached / $2.50 cache write / $10.00 output per 1M tokens vs GPT-5.6 Sol $4.00 / $0.40 / $5.00 / $20.00

• Ultrafast rate — GPT-6.1 Sol $12.00 / $0.60 / $15.00 / $60.00, published via the 6x-Standard rule and its own row in the Ultrafast pricing table vs GPT-5.6 Sol, no published rate

• Long context above 272K input tokens — GPT-6.1 Sol repriced at 2x input and cache rates and 1.5x output for the whole request, giving an Ultrafast long-context line of $24.00 / $1.20 / $30.00 / $90.00 vs GPT-5.6 Sol's Ultrafast equivalent, which nobody has published

• Availability of the tier — GPT-6.1 Sol: all API users, subject to separate rate limits, and Codex plus ChatGPT Work on Pro $500 and eligible Enterprise and Edu plans vs GPT-5.6 Sol: limited preview for select API customers

• Rate limits at launch — GPT-6.1 Sol has its own Ultrafast budget, separate from the Standard and Fast limits, set per organization and raised on request vs GPT-5.6 Sol, not published

• Data residency — GPT-6.1 Sol supports US and EU residency and global processing under Fast and Ultrafast vs GPT-5.6 Sol, US residency and global processing

• Speed multiple — GPT-6.1 Sol, none published vs GPT-5.6 Sol, "up to 14x" in the August announcement, unreproduced

• Context and output ceiling — 1,050,000 and 128,000 for both

One asymmetry in that list is worth pulling out. GPT-6.1 Sol is the only Ultrafast configuration with EU data residency: Astra's Ultrafast is US-only, and 5.6 Sol is not open at all. If your workload is pinned to EU processing, October 8 is the first date on which this tier was available to you in any form.

On a subscription, the difference is a plan, not a model

In Codex and ChatGPT Work both generations behave the same way structurally — Ultrafast is a speed setting on a model you already have — but the gate in front of it is new. Ultrafast is available on Pro $500 and on eligible Enterprise and Edu plans, and it is off by default in Enterprise workspaces until an administrator enables it. Inside a subscription, included usage is consumed at 8x the Standard rate and purchased credits or Enterprise pay-as-you-go are billed at 6x; on the API the multiplier is the flat 6x on every pricing line.

For GPT-5.6 Sol specifically, none of that is reachable yet. There is no subscription path — the ChatGPT-side speed documentation lists Ultrafast as supporting GPT-6 Astra and GPT-6.1 Sol — and the API path is still gated behind a preview cohort. Whatever the 14x figure promises, it is promising it to a set of accounts that has not widened in two months.

A screenshot of OpenAI's API pricing page with the Ultrafast tab selected, showing exactly two rows: gpt-6-astra at $60.00 input, $6.00 cached input, $75.00 cache writes and $300.00 output per 1M tokens short-context, stepping to $120.00 / $12.00 / $150.00 / $450.00 above 272,000 input tokens, and gpt-6.1-sol at $12.00 / $0.60 / $15.00 / $60.00 short-context stepping to $24.00 / $1.20 / $30.00 / $90.00, with the page's note that short context means up to 272K input tokens.

Which one to actually use, and when to wait

If you are starting today, the question mostly answers itself: GPT-6.1 Sol Ultrafast is purchasable, priced, and documented, and GPT-5.6 Sol Ultrafast is not. The reason to reach for the older model instead is not the tier at all — it is that GPT-5.6 Sol has been in the wild since July 9, 2026, and if you have prompt suites, evaluations and cost models built against it, the migration cost of moving to a Sol refresh with a different knowledge cutoff and a strictly narrower reasoning ladder is real, even when the price is halved. Do that move on its own schedule, not because a speed setting finally arrived.

The decision rule for the tier itself is the one Astra's numbers imply and Sol's validate: six times the price for a latency cut is a good trade when a human or an agent is blocked on the output and a bad one when a job has until morning. Nothing about the GPT-6.1 Sol rollout changes that arithmetic; it just makes it available on a cheaper base model, where the absolute bill for a fast turn is smaller.

And if you want the Fast tier instead, you can reason about it from the rate card alone: 2x Standard, and Fast has been the steady middle option since O​penAI renamed Priority processing on July 30, 2026. Ultrafast is still the tier where you are buying an unmeasured promise, and on the GPT-5.6 Sol side, one you cannot buy at all.

Calling either of them from one key

Standard-tier GPT-5.6 Sol and GPT-6.1 Sol are both on OrcaRouter — openai/gpt-5.6-sol at $4.00 input and $20.00 output per million tokens, openai/gpt-6.1-sol at $2.00 and $10.00 — served at 0% markup, with the provider's list price passed through so a vendor price move reaches your bill the same day it reaches the rate card. That matters more than usual on this pair, because GPT-5.6 Sol's $4.00 / $20.00 is explicitly promotional and O​penAI states it is available at least through November 21, 2026. When that window closes, the gap between the two generations narrows, and a pass-through platform means you find out in your usage numbers rather than in a renewal email.

The same key and endpoint carry more than 200 models, so comparing the two Sol generations on your own traffic does not require two integrations or two contracts, and automatic failover keeps a production path standing if one provider degrades. Where a single answer is not enough, the routing DSL composes several models into one call and model fusion puts a panel behind it.

What we do not serve is Ultrafast. On either model it is a service-tier flag on an O​penAI account rather than a separately routable model, so it stays inside your direct O​penAI integration — and on GPT-5.6 Sol it is not even generally available there. We say that plainly because the alternative is implying a capability we do not have. Which is not a sentence we want on this blog.

A screenshot of the OrcaRouter model page for GPT-5.6 Sol, model id openai/gpt-5.6-sol, showing a 1,050,000-token context window, 128K maximum output, text and image input with text output, public benchmarks attributed to OpenAI dated 2026-07-09, input price $4.00 and output price $20.00 per 1M tokens, p50 time to first token 1.36 s, a 2.86 s figure and 85.0M tokens of traffic, with the page's code sample and EN language toggle visible in the header.

Compared in this article1

Detected from this article · Benchmarks: Artificial Analysis · updated daily