
تسعير GPT-5.6 بعد التخفيض: Luna مقابل Terra مقابل Sol، وأي مستوى يجب استخدامه
- metaجديدMeta: Muse Spark 1.22026-08-0557الذكاء72البرمجة
- qwenجديدQwen: Qwen3.8 Max2026-08-0358الذكاء72البرمجة
- deepseekجديدDeepSeek: DeepSeek V4 Flash 07312026-07-3152الذكاء69البرمجة
- minimaxجديدMiniMax: MiniMax-H32026-07-31minimax/minimax-h3
- qwenQwen: Qwen3.7 Flash2026-07-27$0.03 / $0.13 لكل مليون رمز · 2275 tok/s
- orcaOrcaDub: OrcaDub 1.02026-07-27orca/dub
- anthropicAnthropic: Claude Opus 52026-07-2463الذكاء78البرمجة
- googleGoogle: Gemini 3.6 Flash2026-07-2152الذكاء69البرمجة
- googleGoogle: Gemini 3.5 Flash-Lite2026-07-2137الذكاء49البرمجة
- metaMeta: Muse Spark 1.12026-07-1653الذكاء71البرمجة
- kimiMoonshotAI: Kimi K32026-07-1560الذكاء76البرمجة
- openaiOpenAI: GPT-5.6 Luna2026-07-0952الذكاء71البرمجة
- openaiOpenAI: GPT-5.6 Terra2026-07-0957الذكاء77البرمجة
- openaiOpenAI: GPT-5.6 Sol2026-07-0961الذكاء77البرمجة
- grokxAI: Grok 4.52026-07-0856الذكاء72البرمجة
- tencentTencent: Hy32026-07-0642الذكاء59البرمجة
- obsidianQwen3.6 35B A3B Uncensored (Aggressive)2026-07-0232الذكاء42البرمجة
- obsidianGemma4 26B A4B Uncensored (Balanced)2026-07-0226الذكاء39البرمجة
- anthropicAnthropic: Claude Sonnet 52026-06-3055الذكاء72البرمجة
- klingKling: Kling 3.0 Turbo2026-06-1757الذكاء52البرمجة57الرياضيات
After OpenAI's July 30, 2026 price cut, the GPT-5.6 family — GPT-5.6 Luna, GPT-5.6 Terra, and GPT-5.6 Sol — has never been cheaper to run at scale. A week later, on August 6, OpenAI pushed the same tiers into ChatGPT itself: GPT-5.6 Luna is becoming the default model for free and Go accounts with unlimited text chats, while GPT-5.6 Sol was updated for Plus and Pro. But the three tiers are priced very differently, and picking the right one (plus using caching, batch, and routing well) is the single biggest lever on your API bill. This guide lays out the full post-cut pricing, worked cost examples, how each tier compares to rivals, and how to route between them through one endpoint.
Prices are per million tokens (input / output), reflecting the post-cut rate card as reported and OrcaRouter's pass-through model pages; tiered and cache rates come from OrcaRouter's model pages, and competitor figures are attributed. The ChatGPT rollout details in this piece are as OpenAI stated them on August 6, 2026 — vendor-reported, not independently measured. Prices change — verify before building.
ملخص
GPT-5.6 Luna ($0.20 / $1.20) is the high-volume workhorse; Terra ($2 / $12) is the balanced middle; Sol ($5 / $30) is the flagship for the hardest reasoning and agentic coding. All three share a ~1.05M-token context and 128K max output. Prompt caching and the batch option cut costs further, while very long inputs move you to a higher tier. On the consumer side, OpenAI announced August 6 that GPT-5.6 Luna is becoming the default for Free and Go ChatGPT accounts with unlimited text-only chats, and that an updated GPT-5.6 Sol is rolling out to Plus and Pro. Use Luna for the easy majority, Terra when you need more, and Sol only where it earns its cost — and route by difficulty through one OpenAI-compatible endpoint to minimize spend.
النقاط الرئيسية
• لونا: 0.20 دولار / 1.20 دولار لكل مليون توكن — سريع، رخيص، للعمل عالي الحجم والحساس لزمن الاستجابة.
• تيرا: $2 / $12 — الفئة المتوسطة المتوازنة للمهام الأصعب التي لا تحتاج إلى الرائد.
• Sol: $5 / $30 — الرائد في الاستدلال العميق، والبرمجة واسعة النطاق، والوكلاء طويلة المدى.
التخزين المؤقت والمعالجة الدفعية يخفضان التكاليف أكثر؛ والمدخلات الطويلة تنقلك إلى مستوى سياق طويل أغلى ثمنًا.
• ChatGPT (OpenAI-stated, Aug 6): Luna becomes the free/Go default with unlimited text chats; Sol gets a reliability update for Plus/Pro.
• أكبر رافعة توفير: وجّه حسب الصعوبة — لا تدفع أسعار Sol مقابل عمل يمكن لـ Luna إنجازه.
The new GPT-5.6 base pricing
Here's the post-cut base pricing per million input/output tokens: Luna $0.20 / $1.20 (down from $1 / $6), Terra $2 / $12 (down from $2.50 / $15), and Sol $5 / $30 (unchanged). The two cheaper tiers were cut on July 30, 2026; the flagship held. Put simply, the low tiers are now priced for scale, while the flagship stays premium.

التسعير المتدرج والمخزّن مؤقتًا والدفعي
The base rate is only the starting point. GPT-5.6 pricing has three modifiers that materially change your effective cost:
• مستويات السياق الطويل. ترتفع الأسعار للمدخلات الطويلة جدًا. في صفحات التمرير المباشر الخاصة بـ OrcaRouter، يكون سعر Luna $0.20 / $1.20 حتى مستوى سياق كبير و$0.40 / $1.80 بعد ذلك؛ بينما يكون سعر Sol $5 / $30 في المستوى الأساسي و$10 / $45 في المستوى الأكبر. يتم اختيار المستوى بناءً على عدد توكنات الإدخال لكل طلب، لذا فإن بعض المطالبات الضخمة يمكن أن ترفع متوسط سعرك بهدوء.
{{1}}التخزين المؤقت للسياق.{{/1}} {{2}}يتم احتساب السياق المُعاد استخدامه بخصم كبير — قراءة ذاكرة التخزين المؤقت في Luna تبلغ حوالي 0.02 دولار لكل مليون رمز (مقابل 0.20 دولار للسياق الجديد)، مع كتابة ذاكرة التخزين المؤقت بحوالي 0.25 دولار؛ قراءة ذاكرة التخزين المؤقت في Sol تبلغ حوالي 0.50 دولار (مقابل 5 دولارات للسياق الجديد).{{/2}} {{3}}بالنسبة للوكلاء والدردشة مع موجه نظام مستقر، غالبًا ما يكون التخزين المؤقت هو أكبر توفير منفرد.{{/3}}
• المعالجة بالدفعات: تُنفَّذ المهام غير العاجلة بشكل غير متزامن بسعر مخفّض — مثالية للتصنيف الجماعي، والتقييمات، والتوليد دون اتصال حيث لا يهم زمن الاستجابة.
أمثلة عملية: ما هي التكلفة الفعلية
الأرقام الملموسة تجعل المستويات حقيقية. خذ 10 ملايين توكن شهريًا مع تقسيم 70% إدخال / 30% إخراج (مزيج نموذجي من الدردشة/الوكيل):
• لونا: حوالي 5 دولارات شهريًا (حوالي 4.37 دولار مع التخزين المؤقت للطلبات).
• Terra: حوالي 50 دولارًا شهريًا.
• Sol: حوالي 125 دولارًا شهريًا (ما يقارب 109 دولارات مع التخزين المؤقت).
هذا فارق 25 ضعفًا بين Luna وSol لنفس الحجم — وهو بالضبط سبب أهمية مطابقة كل طلب مع الطبقة الأقل تكلفة القادرة على تنفيذه أكثر من أي سعر منفرد. (تتطابق هذه الأرقام مع حاسبة التكلفة الموجودة في صفحة OrcaRouter، والتي تقدّر بناءً على سعر القائمة؛ أرقامك الفعلية تعتمد على التخزين المؤقت ومزيج المدخلات/المخرجات لديك.)
ما هو الغرض الفعلي من كل مستوى
GPT-5.6 Luna — حصان العمل للكميات
Luna is the fast, cost-efficient tier, tuned for high-volume, latency-sensitive workloads: chat, classification, extraction, routing, and lightweight agentic tasks, with a p50 time-to-first-token around 1.65 seconds. After the cut, at $0.20 / $1.20 it's priced to compete directly with cheap open models — the default for the easy majority of calls. It is also the tier OpenAI is steering consumer ChatGPT toward: per its August 6 announcement, GPT-5.6 Luna is becoming the default model for Free and Go accounts, with text-only chats going unlimited and a new Think button for higher reasoning on harder questions arriving the following week (file, image, and voice limits remain). OpenAI says Luna makes 62% fewer factual errors than the GPT-5.5 Instant it replaces — a vendor-reported figure, but a clear signal that the volume tier is where the company is pointing most of its traffic.
GPT-5.6 Terra — الوسط المتوازن
تيرا تقع بين الحجم والرائد: أكثر قدرة من لونا للتفكير والبرمجة الأكثر صعوبة، ولكنها أرخص بكثير من سول. بسعر 2$ / 12$ فهي الخيار الافتراضي المعقول عندما لا تكون لونا كافية تمامًا ولكنك لا تحتاج إلى الرائد — مهام الاستخراج متوسطة التعقيد، والمسودات، والمهام متعددة الخطوات التي لا تزال تعمل على نطاق واسع.
GPT-5.6 Sol — الطراز الرائد
Sol is built for the hardest work: deep multi-step reasoning, large-scale software engineering, and long-horizon agentic workflows, staying coherent across a ~1.05M-token context and up to 128K output. At $5 / $30 (base) it's a premium choice — reserve it for tasks that genuinely need it, like complex multi-file coding or long agent runs. On August 6, OpenAI also updated GPT-5.6 Sol in ChatGPT for Plus and Pro users: the company says the chat version is more reliable with facts — 68% fewer factual errors, per OpenAI — and gives more focused answers, with a new slider to control how much reasoning effort it applies. The update is chat-only (the Sol behind Work and Codex is unchanged), and the API tier stays at $5 / $30.
تحذير بشأن رمز الاستدلال
GPT-5.6 are reasoning models, so effective output cost can exceed a naive estimate: when reasoning is on, internal reasoning tokens are billed as output. On hard prompts with high reasoning effort, that can dominate your bill. Tune reasoning effort to the task (low or off for simple calls), cap output tokens where you can, and measure actual usage. Cheaper per-token rates help; generating fewer tokens helps more.
كيف تقارن كل فئة بالمنافسين
The cut repositioned GPT-5.6 against the field. Per reporting, Luna's $0.20 / $1.20 now undercuts Claude Haiku 4.5 by roughly 5x on input and 4x on output, and Terra's $2 / $12 falls below Claude Sonnet 5's standard pricing (reported at $3 / $15). Against open weights, Luna sits near the floor set by DeepSeek's V4 line and Zhipu's GLM-5.2 (about $1.20 / $4.10), with Qwen and Gemini Flash tiers nearby. The upshot: for cost-sensitive work, Luna is now competitive with the cheapest capable models rather than a premium alternative to them — while Sol remains a genuine premium tier for capability you can't get cheaply.
أداة الادخار الحقيقية: التوجيه حسب الصعوبة
The cheapest bill isn't a single tier — it's matching each request to the least expensive model that can do it. In practice: send easy, high-volume calls to Luna (or a cheap open model), step up to Terra for harder tasks, and use Sol only for the genuinely difficult minority. Combined with caching and batch, this routinely cuts costs far more than any single price change. The catch is operational: you don't want to re-integrate three OpenAI tiers plus open-model alternatives separately.
الوصول إلى الثلاثة جميعها (والبدائل الأرخص) من خلال نقطة نهاية واحدة.
This is where a vendor-neutral router helps. OrcaRouter exposes GPT-5.6 Luna, Terra, and Sol — at the same post-cut prices, 0% markup — through one OpenAI-compatible endpoint, alongside cheaper open models like DeepSeek V4 Pro, GLM-5.2, and Qwen. Switching tiers (or A/B testing Luna against an open model) is a config change, not a re-integration, and you can route each request to whatever is cheapest.

The post-cut Luna rate card is visible on OrcaRouter's own model page at list price — $0.20 / $1.20 per million tokens — because the router passes provider prices through at 0% markup, with an on-page cost calculator for a typical monthly bill. There's also a free tier and an Offers page for additional savings.

الأسئلة الشائعة
What are the GPT-5.6 prices after the cut?
لونا 0.20 دولار / 1.20 دولار، وتيرا 2 دولار / 12 دولارًا، وسول 5 دولارات / 30 دولارًا لكل مليون رمز إدخال/إخراج. تم تخفيض لونا وتيرا في 30 يوليو 2026؛ بينما سول لم يتغير.
Which GPT-5.6 tier should I use?
Luna للأعمال عالية الحجم والحساسة لزمن الاستجابة؛ Terra للمهام الأصعب التي لا تحتاج إلى النموذج الرئيسي؛ Sol لأصعب مهام التفكير والبرمجة بالوكلاء. وجّه حسب الصعوبة لتقليل التكلفة.
Is GPT-5.6 Luna free in ChatGPT?
OpenAI announced on August 6 that GPT-5.6 Luna is becoming the default model for free and Go ChatGPT accounts, with unlimited text-only chats; limits remain on file uploads, images, and voice tools, and a Think button for higher reasoning is rolling out separately. These are OpenAI-stated plans, not independently verified.
How much does GPT-5.6 cost per month?
عند 10 ملايين رمز شهريًا (70% مدخلات): حوالي 5 دولارات على لونا، و50 دولارًا على تيرا، و125 دولارًا على سول — فرق 25 ضعفًا، قبل التخزين المؤقت. تعتمد تكلفتك الفعلية على التخزين المؤقت ومزيج المدخلات/المخرجات.
هل جميع المستويات الثلاثة لها نفس نافذة السياق؟
{{1}}نعم — سياق يبلغ حوالي 1.05M رمزًا وما يصل إلى 128K رمزًا للإخراج، مع دعم الرؤية والأدوات وJSON والاستدلال.{{/1}}
كيف يؤثر التخزين المؤقت للمطالبات على التكلفة؟
يقلل التخزين المؤقت بشكل كبير من تكلفة السياق المتكرر {{1}}— قراءة ذاكرة التخزين المؤقت في Luna تبلغ حوالي 0.02 دولار لكل مليون رمز مقابل 0.20 دولار للسياق الجديد {{/1}}— لذا أعد استخدام المطالبات المخزنة مؤقتًا حيثما أمكن. {{2}}وعلى العكس، يمكن أن تنقلك المدخلات الطويلة إلى فئة سياق طويل أغلى ثمنًا.{{/2}}
كيف تقارن المستويات بـ Anthropic أو النماذج المفتوحة؟
Per reporting, Luna undercuts Claude Haiku 4.5 (~5x input / 4x output) and Terra falls below Claude Sonnet 5 ($3 / $15). Luna is also near the cheap open-model floor (DeepSeek, GLM-5.2 ~$1.20/$4.10).
أين يمكنني استخدام المستويات الثلاثة معًا؟
Through OrcaRouter's single OpenAI-compatible endpoint at 0% markup, alongside cheaper open models for difficulty-based routing.
خلاصة القول
After the cut, GPT-5.6 Luna ($0.20 / $1.20) and Terra ($2 / $12) are compelling for high-volume and mid-tier work, while Sol ($5 / $30) remains the premium flagship — a 25x cost spread that rewards smart routing. The consumer rollout reinforces the split: OpenAI is making Luna the free default while reserving Sol's reliability update for paid plans, which for API builders is one more reason to route easy traffic to Luna. The biggest savings come from routing by difficulty, leaning on caching and batch, and controlling reasoning tokens rather than defaulting to one tier. Doing that is easiest through a single 0%-markup endpoint like OrcaRouter, where all three tiers (at the new prices) sit next to the cheaper open models you'll want to compare them against.
مقارنات في هذه المقالة2
مستخرج من هذه المقالة · المعايير: Artificial Analysis · يُحدَّث يوميًا
