'세 개의 게시물 깊이'라는 제목의 생성된 히어로 카드로, 부제는 'GPT-6.1 Ultrafast는 ChatGPT에서만 Pro $500 — 모든 API 사용자에게 개방'이고, 칩에는 'Pro $500 요금제', 'Enterprise 및 Edu 대상', '모든 API 사용자'라고 적혀 있으며, 바닥글에는 'OpenAI 문서 기준 제공 여부, 2026년 10월 8일'이라고 적혀 있습니다.
Guides & Insights

GPT-6.1 Ultrafast는 Pro, 단 $500뿐이며, 그 세부 내용은 게시물 세 개 깊숙이 묻혀 있었다

작성자

Alistair Wren

게시일

최신 모델 · 20모든 모델 보기 →
벤치마크: Artificial Analysis · 매일 업데이트
모든 게시물로 돌아가기

GPT-6.1 Sol Ultrafast arrived on October 8, 2026 with the rarest thing a speed tier can ship with: a published price. $12.00 per million input tokens, $60.00 per million output tokens, exactly six times what GPT-6.1 Sol — the model O​penAI released on September 29, 2026 at its DevDay 2026 keynote — costs on the standard lane. What the announcement thread did not lead with is where a ChatGPT subscriber can actually switch it on, and the answer, three posts further down, is the $500 Pro plan. That ordering is worth more attention than it got, because the API half and the subscription half of this rollout are gated in opposite ways, and the difference decides whether Ultrafast is something you can try this week or something you have to buy a plan for.

모델에 관해 바뀐 것은 아무것도 없습니다. GPT-6.1 Sol Ultrafast는 동일한 체크포인트, 동일한 1,050,000토큰 컨텍스트 윈도, 동일한 128,000토큰 출력 상한, 동일한 2026년 4월 30일 지식 컷오프, 그리고 GPT-6.1 Sol과 동일한 답변을 제공합니다. Ultrafast는 서비스 티어, 즉 요청하는 스케줄링 우선순위이지 두 번째 모델이 아닙니다. 제품의 전부는 지연 시간이며, 몇 배의 가격으로 판매됩니다. 그리고 중요한 두 가지 질문은 여러분이 그것을 얼마나 얻느냐와 누가 그것을 구매할 수 있느냐뿐입니다.

기록이 실제로 말하는 것, 말하는 순서 그대로

The grievance version of this story is that O​penAI buried the gate. The accurate version is more specific, and it is checkable line by line against O​penAI's own documentation.

개발자 체인지로그에는 10월 8일자 항목이 실려 있다: Responses API에서 GPT-6.1 Sol용 Ultrafast 모드를 추가하며, 요청하려면 model: "gpt-6.1-sol"과 함께service_tier: "ultrafast"를 보내면 된다. 이는 생성된 출력 토큰 사이의 시간을 줄이는 것으로 설명되고, 요청 속도 제한의 적용을 받는 모든 API 사용자가 사용할 수 있으며, 글로벌 처리와 미국 및 EU 데이터 레지던시를 제공한다고 명시되어 있다. 그 자체만 놓고 보면, 그것은 공개된 요금으로의 출시다.

등급 문서에 따르면 Ultrafast는 API에서 가장 빠른 서비스 등급으로, GPT-6 Astra와 GPT-6.1 Sol에서 폭넓게 이용할 수 있으며 GPT-5.6 Sol에서는 프리뷰로 접근할 수 있고, 이를 위해 WebSockets를 권장한다. 여전히 요금제에 대한 언급은 없다.

A screenshot of OpenAI's Ultrafast mode documentation page inside the API production guide, headed 'Ultrafast mode' and described as the fastest API service tier for latency-sensitive work, stating it is broadly available for GPT-6 Astra and GPT-6.1 Sol with preview access for GPT-5.6 Sol.

구독 세부 정보는 ChatGPT 측 속도 페이지에 있으며, 한 문단에 세 가지 제한이 겹쳐 있습니다: Ultrafast는 Codex와 ChatGPT Work에서, Pro $500에서, 그리고 자격을 갖춘 Enterprise 및 Edu 요금제에서 사용할 수 있습니다. Enterprise 관리자는 기본적으로 꺼진 상태로 제공되며 사용자별 또는 워크스페이스별로 직접 활성화해야 합니다. 그리고 소비자용 Chat 인터페이스에는 이 등급이 전혀 제공되지 않습니다 — GPT-6.1 Sol 자체가 Work와 Codex에 있으므로, Chat은 애초에 범위에 포함되지 않았습니다.

Put the three pages together and the shape is clear. An API key is a ticket. A ChatGPT subscription is a ticket only if it is the most expensive one O​penAI sells, or if an administrator has enabled it for you. Nobody reading the first post of the announcement would guess that.

게이트가 존재하는 이유, 그리고 그것은 임의적이지 않습니다.

The obvious reading is that O​penAI is using the fast lane to sell $500 plans. The mechanism underneath is less cynical and more useful: Ultrafast is an expensive meter, and a $500 plan is the only consumer plan where an 8x meter fits inside the included usage.

구독에서는 Ultrafast를 사용할 때 포함된 사용량이 표준 요율의 8배로 소모되므로, Ultrafast 턴 한 번은 표준 턴 여덟 번과 같은 허용량을 소진합니다. 이는 해당 등급이 하는 일의 결과입니다 — 서빙 용량을 소비해 실제 시간을 확보하는 것이고, 구독은 고정 용량 상품이기 때문입니다. 용량 소비가 8배인데 같은 허용량을 청구하면 월 $20 요금제의 경제성은 무너집니다. 허용량을 8배로 청구하면 그 등급은 소규모 요금제에서 그저 매력이 없어집니다. 어느 쪽이든, 그 등급은 허용량이 이를 흡수할 만큼 충분히 큰 곳에 자리하게 됩니다.

The API has no such constraint, because it is metered. There, Ultrafast gets its own rate-limit budget, separate from the Standard and Fast budgets, and O​penAI describes those limits as set per organization rather than published on the tier page. That is a second thing to watch that the announcement did not headline: switching a workload to Ultrafast does not just multiply the bill, it moves the workload onto a different ceiling, and if you raise traffic beyond what the organization was granted, the tier's own limits become the failure mode.

두 가지 방식으로 속도를 사는 산수

숫자로 정리한 결정은 다음과 같으며, 이 부분은 제 토큰 수가 아니라 여러분 자신의 토큰 수로 따져볼 가치가 있습니다. 한 턴에 입력 토큰 30,000개를 보내고 출력 토큰 1,500개를 받는 에이전트 세션이 40턴 동안 이어진다고 해보세요.

• Total tokens per run — 1,200,000 input, 60,000 output, spread over 40 requests that each sit comfortably under the 272,000-token threshold where O​penAI reprices the whole request.

• 표준 GPT-6.1 Sol 기준 — 120만 토큰이 백만 개당 $2.00이면 $2.40이고, 여기에 60,000개가 백만 개당 $10.00이면 $0.60입니다. 실행당 3달러입니다.

• GPT-6.1 Sol Ultrafast의 경우 — 120만 개에 $14.40, 추가로 6만 개는 $60.00에서 $3.60. 실행당 18달러입니다.

• 차이점 — 그 외에는 달라지지 않는 작업에서 생성 시간의 대부분을 없애는 데 드는 실행당 $15.00.

실행당 $15.00은 구독과 견줘 따져야 할 숫자입니다. Pro 요금제를 고려하는 유일한 이유가 Ultrafast라면, 이 토큰 수 기준으로 한 달에 약 33회 실행하는 지점이 요금제의 $500와 API의 토큰당 종량 과금이 교차하는 지점입니다 — 그보다 적으면 같은 등급을 구매하는 더 저렴한 방법은 종량 과금이고, 그보다 많으면 요금제가 이기기 시작합니다. 토큰 수에 따라 이 교차점은 달라지며, 때로는 크게 달라지므로 33을 임계값으로 삼기 전에 직접 숫자를 계산해 보세요. 이는 명시된 작업량을 바탕으로 한 우리의 계산이며, 공식 발표된 손익분기점이 아닙니다.

A generated cost card headed 'The same 40-turn agent, two bills', showing two panels: GPT-6.1 Sol on the standard tier at $3.00 per run and GPT-6.1 Ultrafast at $18.00 per run, both for 1,200,000 input and 60,000 output tokens, with a footer reading that prices are OpenAI's own per-million list rates and the arithmetic is ours.

p>어떤 토큰 수에서도 살아남는 두 가지 구조적 메모.

속도 주장은 빌려온 것이며, 이것이 바로 주의 깊게 읽어야 할 단 하나의 부분이다.

Ultrafast is sold on "up to 8x faster", and the only model-specific measurement O​penAI publishes behind that phrase belongs to a different model. The sentence in its documentation reads that GPT-6 Astra Ultrafast generates tokens up to 8x faster than GPT-6 Astra in Standard mode in Codex. That is an Astra figure, measured in Codex. The documentation for GPT-6.1 Sol's tier describes it as reducing inter-token time and points at a rate card; it does not put a number on it, and no independent party has published a tokens-per-second measurement for the Sol variant either.

그 주장은 충분히 성립할 수 있다. 두 모델 모두 같은 서빙 스택에서 실행되고 티어 메커니즘도 동일하므로, Astra 배수는 Sol에 대한 합리적인 사전값이다. 하지만 "최대 8배"는 벤더가 제시한 상한이며, 재현된 적이 없고, 형제 모델에서 가져온 것이다 — 그리고 그 문장에서 "최대"라는 말은 실제로 중요한 역할을 한다. 그 배수는 자신의 프롬프트에서 직접 테스트할 가설로 취급하라. 왜냐하면 해당 티어의 자체 문서는 결코 당신을 대신해 그것을 테스트해 주지 않기 때문이다.

같은 유보 사항은 더 오래된 티어에 대해서는 훨씬 부담 없이 언급할 수 있다. GPT-5.6 Sol을 앞선다는 Ultrafast는 2026년 8월 13일 "Standard 처리보다 최대 14배 빠름"을 내세우며 제한적 프리뷰로 발표되었고, 세 달이 지난 지금도 문서에는 여전히 프리뷰 액세스라고 적혀 있는 반면 Ultrafast 요금표에는 정확히 두 개의 행, 즉 GPT-6 Astra와 GPT-6.1 Sol만 실려 있다. 공개된 가격도, 일반 제공도 없는 14배라는 숫자는 제품이라기보다 주장에 가까우며, 이 티어 사다리의 얼마나 많은 부분이 아직 희망 사항에 불과한지를 잘 일깨워 준다.

$500 플랜이 진행되지 않는다면

지연 시간 티어를 테스트하려고 Pro $500를 살 생각이 없는 사람이라면, 유용한 방법은 이번 발표가 묶어 놓은 두 가지를 분리해 보는 것입니다. 이 티어는 API에서 모든 API 사용자에게 제공되므로, 패스트 레인은 구독 없이도 구매할 수 있습니다 — 백만 토큰당 $12.00 및 $60.00에, 자체 레이트 리밋 예산을 갖추고, 측정할 수 있을 만큼 작은 워크로드에서요. 먼저 그곳에서, 자신의 프롬프트로, 표준 레인과 비교해 테스트해 보세요.

The standard lane is the one that exists without any of this. OrcaRouter serves GPT-6.1 Sol as openai/gpt-6.1-sol at O​penAI's own list rates — $2.00 per million input tokens and $10.00 per million output tokens — with 0% markup and the provider's price passed straight through, so a vendor repricing lands on our side the same day. Ultrafast is not something we sell: it is a service-tier flag billed on your own O​penAI account, and we would rather say that plainly than let a model page imply otherwise. What one key does buy is the ability to route the standard lane and the rest of the catalogue — more than 200 models behind one O​penAI-compatible endpoint — and to fail over automatically across providers when one is degraded, which is the cheapest insurance available if you are about to put a latency tier in front of a production path.

A generated availability card headed 'Where GPT-6.1 Ultrafast can be switched on', with an 'Included' column listing the API for all users subject to rate limits, Codex and ChatGPT Work on Pro $500, and eligible Enterprise and Edu with administrator opt-in, and a 'Not included' column listing ChatGPT Plus and Team plans and the consumer Chat surface, footnoted to OpenAI's API changelog dated 8 October 2026 and the ChatGPT speed page.

그리고 Plus나 Team 요금제를 쓰고 계시면서 찾고 있던 답이 "이거 그냥 켜면 되나요?"였다면, 답은 아니오이며, 구체적인 조건이 있습니다. Codex와 ChatGPT Work는 Pro $500 또는 자격을 갖춘 Enterprise나 Edu 요금제에서 사용할 수 있고, Enterprise 종류라면 관리자가 활성화해야 합니다. 그 외의 모든 것은 API를 통해 Ultrafast에 도달하며 토큰당 비용을 지불합니다.

다음에 볼 콘텐츠

세 가지가 있다면 이것을 제한적 출시에서 평범한 출시로 바꿔 놓을 것이며, 그 세 가지 모두 관찰 가능하다. GPT-6.1 Sol에 대해 공개된 Ultrafast 요청 한도는 해당 티어 자체의 상한이 청구서보다 먼저 걸리는지에 대한 추측을 끝내 줄 것이다. 모델별 속도 측정 — 또는 독립적인 측정 — 은 차용한 Astra 수치를 Sol 티어를 평가할 수 있는 무언가로 대체할 것이다. 그리고 Ultrafast가 ChatGPT Plus에 도달한다면 용량 계산이 바뀌었다는 뜻이며, 이는 게이트가 티어의 영구적인 형태라기보다 출시 초기의 희소성 때문이었다는 신호다.

그때까지 정직한 요약은 좁고 검증 가능하다. GPT-6.1 Sol은 2026년 9월 29일에 출시되었고 변경되지 않았다. 2026년 10월 8일에는 실제 가격이 붙은, API 사용자에게는 개방되고 ChatGPT에서는 $500 플랜 뒤에 잠긴 패스트 레인이 생겼다 — 이는 세 번째 글이 아니라 첫 번째 글에 들어갔어야 할 세부 사항이며, 대부분의 독자에게 이것이 개별 항목인지 구매 주문인지를 결정하는 세부 사항이다.

이 글에서 비교한 모델1

이 글에서 자동 인식 · 벤치마크: Artificial Analysis · 매일 업데이트