
Atria Dawn 対 GPT-5.6 Sol:13ポイントの指標差と、それを無視する研究ループ
- deepseekNEWDeepSeek: DeepSeek V4.1 Flash2026-09-1040知能
- openaiNEWOpenAI: GPT-6 Astra2026-09-0453知能77コーディング
- googleNEWGoogle: Gemini 3.8 Flash2026-09-0241知能76コーディング
- qwenNEWQwen: Qwen3.8 Max (0902)2026-09-0240知能72コーディング
- anthropicNEWAnthropic: Claude Fable 5.12026-09-0153知能82コーディング
- AlibabaQwen: Qwen3.8 Flash2026-08-26$0.15 / $0.47 100万トークンあたり
- z-aiZ.ai: GLM 5.3 Flash2026-08-2642知能72コーディング
- DeepSeekDeepSeek: DeepSeek V4 Flash Vision (Exp)2026-08-21$0.22 / $0.66 100万トークンあたり
- z-aiZ.ai: GLM 5.32026-08-1845知能75コーディング
- obsidianQwen3.8 27B2026-08-1534知能68コーディング
- deepseekDeepSeek: DeepSeek V4 Pro 08132026-08-1236知能69コーディング
- grokSpaceXAI: Grok 4.62026-08-1244知能77コーディング
- metaMeta: Muse Spark 1.22026-08-0540知能72コーディング
- qwenQwen: Qwen3.8 Max2026-08-0340知能72コーディング
- deepseekDeepSeek: DeepSeek V4 Flash 07312026-07-3135知能69コーディング
- minimaxMiniMax: MiniMax-H32026-07-31minimax/minimax-h3
- qwenQwen: Qwen3.7 Flash2026-07-27$0.03 / $0.13 100万トークンあたり
- orcaOrcaDub: OrcaDub 1.02026-07-27orca/dub
- anthropicAnthropic: Claude Opus 52026-07-2451知能78コーディング
- googleGoogle: Gemini 3.6 Flash2026-07-2134知能69コーディング
Here is the entire tension of this comparison in one line: Atria Dawn Preview is a brand-new open-weights agentic model from the Shanghai AI Laboratory, released on September 14 on a 744B MoE GLM-5.2 base with a 256K context and no published price, while GPT-5.6 Sol is OpenAI's flagship tier of the GPT-5.6 family — shipped July 9, priced at $4.00/$20.00 per million tokens, and carrying an independent Artificial Analysis Intelligence Index of 61 that has been stable through two price cuts. On paper the gap is enormous: a 13-point independent index delta between a closed frontier flagship and an unproven open preview. The reason the matchup is worth reading anyway is that Atria Dawn Preview was not built to close that gap — it was built to make it irrelevant, by targeting the one thing a raw index score does not measure: whether a model will keep working until it has a verifiable result, without a human holding the thread.
Sourcing labels carry this whole article, so they come first. GPT-5.6 Sol's index score, context, and pricing are independent facts you can check on Artificial Analysis and OpenAI's own rate page today. Atria Dawn Preview's benchmark rows are vendor-reported from its model card, unreproduced by any neutral lab, and its international API at api.atria-asi.ai has no published price yet. Where the card pits the two against each other, Atria Dawn Preview reports leads on discovery and tool-use suites (BrowseComp 92.5 vs 90.8, BFCL v4 77.0) while GPT-5.6 Sol's coding rows sit far ahead (SWE-bench Pro 59.6 vs a reported 74.7-equivalent on the same grid, Terminal-Bench 2.1 78.3 vs 90.2). None of the Atria numbers is independently verified; treat them as the lab's pitch, not a scorecard.
「model」という語だけを共有する2つのオブジェクト
GPT-5.6 Sol is OpenAI's flagship reasoning tier: a 1.05M-token context window, up to 128K output, text/image/file input, an AA Intelligence Index of 61, and a $4/$20 promotional rate that OpenAI cut from $5/$30 on August 24 and has committed to keep at least through November 21 — with a 272K-token cliff above which the entire request bills at $8/$30. It is the model behind Codex, behind Work, and behind the default ChatGPT experience; it is the most widely deployed reasoning model in production right now. It is also proprietary: you rent it, you never own it.
Atria Dawn Previewは、正反対の経済的対象です。MITライセンスで、重みをダウンロード可能(BF16が353シャード、FP8が177シャード)、テキスト専用で256Kのウィンドウを持ち、SGLangとvLLM経由でデプロイ可能で、同じモデルIDでapi.atria-asi.aiに国際的にホストされています。その売りは生の推論力ではなく、研究ループです。すなわち、分析し、設計し、ツールを使い、コードを書き実行し、結果を読み、復帰し、反復する。同ラボの4つの柱——Discovery、Creation、Delivery、Cybersecurity——はいずれもこのループに帰着します。GPT-5.6 Solはダウンロードできませんが、Atria Dawn Previewはダウンロードして自分が所有するハードウェアで動かせます。そしてその違いこそが、同ラボが賭けているもののすべてです。

• 価格 — Atria Dawn Preview 未公開(国際API)vs GPT-5.6 Sol 100万トークンあたり $4.00/$20.00(入力272K超で $8/$30)
• コンテキスト — Atria 256K テキスト専用 vs GPT-5.6 Sol 1.05M、128K 出力
• 独立記録 — Atria 対 AA Index 61 については該当なし、2回の価格改定を通じて安定
• オープン性 — AtriaはMITの重みでセルフホスト可能、GPT-5.6 SolはプロプライエタリでAPIのみ
インデックスギャップが実際に測定しているもの — そしてそれが見逃しているもの
13ポイントのインデックス差は現実のものであり、ごまかして片付けるべきではない。あなたのワークロードが難しい推論問題——証明、厄介なコードバグ、密度の高い法律文書——であるなら、GPT-5.6 Sol は平均して Atria Dawn Preview を上回り、その分のプレミアムが価格に反映されている。しかし、このインデックスは主に短い単発の推論タスクから構成されており、まさにそこが、リサーチループ型エージェントがルールを変えようとしている領域だ。Atria Dawn Preview が報告している強みは短い回答ではなく、BrowseComp、DeepSearchQA、AutomationBench だ——モデルが何度もターンを重ねて読み、行動し、自己修正できる長期ホライズンのスイートである。ベンダーが報告した内訳(ディスカバリー/ツール使用では勝ち、コーディングでは負ける)は、インデックス向けではなく反復作業向けに調整されたモデルならではの特徴だ。

For a team that already runs GPT-5.6 Sol for multi-step agentic coding, the honest reading is: Atria Dawn Preview is not a replacement, because its coding rows are exactly where it loses. For a team whose pain is the loop itself — the model stopping to ask, losing the thread, never finishing the experiment — the newcomer is aimed at a real gap, and its open weights mean you can find out for yourself whether the loop holds without paying OpenAI's rate for the privilege. Through a routing layer such as OrcaRouter, both models sit behind one key — GPT-5.6 Sol at pass-through list price with automatic failover across providers, and, the moment the Shanghai lab publishes a rate for Atria, the same key routes to it too, so a head-to-head eval costs setup time only.

評決
インデックスのギャップをすべての話として読まないでください。また、ベンダーのグリッドを証拠として読まないでください — 一方の列は独立して検証されており、もう一方はラボの自己申告です。今日、GPT-5.6 Solでソフトウェアを出荷しているなら、ここにあるものは切り替えを主張するものではありません。13ポイントの差分と1.05Mのウィンドウがあなたのワークロードを勝ち取ります。研究や、失敗がモデルの諦めであるような長期ホライゾンの自動化を運用しているなら、Atria Dawn Previewのオープンウェイトとループファーストのチューニングは、制御された評価に値する正当な実験です — 自社のハードウェア上で、フェイルオーバーの背後で、まだ誰も中立にスコアリングしていないプレビューに本番パスを賭けることなく。
