Odyssey-3 與 Wan 3.0 的標題卡,顯示 Odyssey-3 為自 2026 年 10 月 8 日起的研究預覽版,未公佈價格或 API,Physics-IQ Verified 淨提升 +11.15 pp,排名第 02;對比 Wan 3.0 為計量制的 API,自 2026 年 8 月 24 日起全面可用,720p 每秒 0.6 元。
Guides & Insights

Odyssey-3 對決 Wan 3.0:當世界模型遇上影片工廠

作者

Elias Hawthorne

發佈日期

最新模型 · 20查看全部模型 →
基準測試:Artificial Analysis · 每日更新
返回全部文章

Odyssey-3 and Wan 3.0 both get filed under "video model," and that shared label is carrying more weight than it can bear. Odyssey-3 is a learned dynamical system — a simulator that builds an environment from a prompt and predicts how that environment changes as someone moves through it or triggers an event — released by Odyssey on 8 October 2026 as a research preview. Wan 3.0 is a production video generator from A​libaba Cloud, generally available since its 24 August 2026 launch, billed by the second, and already scored on two independent leaderboards. One of these wants to be a place you put a camera. The other wants to be the thing that renders the footage you publish.

Put the two launch pages side by side and the split is immediate. Odyssey leads with physical accuracy — how well its model understands what objects do when they meet. A​libaba leads with throughput, duration and reference control: 30-second takes, native audio, up to ten images and five videos supplied per generation. Neither is a worse version of the other. They are answers to different questions, and the answers change what a team should do this week.

在進入數字之前,先說明一點背景。本文中大多數數據來自廠商自家的發布資料,尚未被這些實驗室以外的任何人重現。若某項數字來自獨立評測機構而非新聞稿頁面,文中會加以說明。這個區別在這裡比平常更為要緊,因為這兩個模型所公開、可供查證的資訊量天差地遠——一個在開放 API 上以秒計費販售,另一個則根本還沒公布價格。

每一個實際上是什麼

Odyssey 對 Odyssey-3 的描述對一篇發布文章而言異常具體,值得直接引用而非概略帶過:它是一套「學習式動力系統,以自迴歸擴散 Transformer 實作,能預測物件如何在空間中移動與互動,以及情境如何隨時間演變」。它回傳的東西並非一般意義上的短片,而是一種在被驅動時持續演變的狀態。該模型推出兩種尺寸——832×480 的 Odyssey-3 與 1280×720 的 Odyssey-3 Pro——Odyssey 表示其預覽支援第一人稱與第三人稱導航,以及獨立攝影機運動。這個版本也包含一個少步數蒸餾變體,透過分佈匹配與對抗蒸餾的組合產生,廠商稱之為即時互動。

Wan 3.0 is a generator on the established model: a prompt and one or more references go in, a finished piece of video comes out. A​libaba's headline capabilities are a 30-second single-run generation, native audio, and reference inputs spanning text, image, video and audio — plus documents in doc, xls, ppt, pdf, md, txt, key, pages and numbers formats, up to 100 MB and 50 pages. Editing is instruction-based, meaning a generation can be modified in visuals, plot or dialogue without regenerating from scratch. A​libaba also states plainly that audio quality and on-screen text accuracy still need work, which is a rarer admission in a launch post than it should be and a useful one to have on the record.

• 輸出內容是什麼——你實際在內操作的模擬環境(Odyssey-3),對比你發佈的算繪短片(Wan 3.0) • 廣告標示尺寸——832×480,或在 Odyssey-3 Pro 上為 1280×720,對比 480p、720p 與 1080p • 單次執行長度——只要環境持續被驅動就可一直執行,對比每次生成 30 秒 • 音訊——在兩種尺寸上都不是主打能力,對比原生音訊但附有供應商標示的品質限制 • 輸入——一段定義環境的提示詞,對比文字、圖像、影片、音訊以及最高 100 MB、50 頁的文件 • 編輯模式——變更狀態並重新模擬,對比對完成後的生成內容進行指令式編輯 • 已記載的 API——未公開,對比自 2026 年 8 月 24 日起開放並計量 • 公佈價格——無,對比 480p / 720p / 1080p 每秒 0.3 / 0.6 / 1.2 元

輸出一秒要付出什麼代價,以及為什麼單位對不上

Wan 3.0 is priced in the most legible unit available to a buyer: time. A​libaba's rates are 0.3 yuan (about $0.05) per second of generated video at 480p, 0.6 yuan (about $0.10) at 720p, and 1.2 yuan (about $0.20) at 1080p. A full 30-second run therefore lands near 9 yuan ($1.50) at 480p, 18 yuan ($3.00) at 720p, or 36 yuan ($6.00) at 1080p. Those are the August launch rates; the launch promotion that cut them by 30% ran only through 23 September, so the numbers above are the ones a buyer meets today. Providers that resell Wan 3.0 typically quote per-minute figures instead, which is worth checking against the per-second arithmetic before any budget is signed off.

Odyssey-3 並沒有可比較的數字,因為 Odyssey 尚未公布價格、費率表、授權條款或 API 文件。取而代之的是第三方基準測試中的一項成本數字,該數字是根據供應商公開陳述的假設計算而來:每 MI355X GPU 小時 1 美元。在此基礎上,Physics-IQ Verified 列出 Odyssey-3 Pro 為每支生成影片 $0.267,Odyssey-3 為 $0.139,兩者都標準化為 24 FPS 與 1280 寬的輸出。同一看板另行指出,透過 Odyssey 的 API 進行提示改寫,估計為每支提交影片 $0.01。請仔細閱讀——這是建模基礎,不是價格。它告訴你的是,在供應商假設的硬體費率下,模擬會花費多少,而不是 Odyssey 實際會開立發票收取的金額。

對買方而言,這兩項事實指向相反的方向。Wan 3.0 是一項按用量計費的明細項目,其最壞情況可以在算繪任何內容之前先計算出來。Odyssey-3 每部影片的模型推估成本較低,卻沒有發票可與之對應——而這正是採購對話會停滯的階段。若要讓比較公平,讀者必須同時掌握兩邊:有真實帳單支撐的真實費率,以及背後尚無任何依據的估計費率。

Two-column scoreboard comparing Odyssey-3 and Wan 3.0 across six dimensions: what you get, advertised sizes, single-run length, published price, documented API, and Physics-IQ Verified. Odyssey-3 returns a simulated environment, ships at 832x480 or 1280x720 on Odyssey-3 Pro, runs as long as it is driven, has no published price or documented API, and scores +11.15 pp at rank 02 on Physics-IQ Verified; Wan 3.0 returns a rendered clip, ships at 480p, 720p and 1080p, renders 30 seconds per run, costs 0.3 / 0.6 / 1.2 yuan a second, has been open and metered since 24 August 2026, and is not listed on that board.

他們共用的那唯一一塊記分板,以及上頭缺少的那個名字

Physics-IQ Verified 是由 Anates Labs 與 DeepMind 發布的動態排名,用來評估影片模型處理物理原理的表現,也是這組配對唯一共有的獨立排行榜——而這正是比較真正變得有趣的地方,因為它同時也是比較瓦解的地方。

Odyssey 的發布貼文指出,Odyssey-3 Pro「在 Physics-IQ Verified 的基準測試上樹立了新的最高水準,並在 WorldMark 的 4 個類別中拿下 3 個第一。」其中 WorldMark 的那一半,在 Odyssey 公布的數字上確實站得住腳:第一人稱風格化(77.2)、第三人稱寫實(79.0)與第三人稱風格化(76.3)名列第一,而第一人稱寫實則以 80.6 位居第二,落後 AlayaWorld 的 83.0 與 Lyra 2.0 的 84.4。這些全都是 Odyssey 自家評測跑分中的廠商自報數字。

「最先進」的那半句,是讀者應該放慢速度的地方。在目前可取得的 Physics-IQ Verified 榜單版本上,若依「高於賽道平均值的淨改善百分點」排名,Black Forest Labs 的 FLUX 3 [large] 以 +12.27 個百分點位居第 01 名。Odyssey-3 Pro 以 +11.15 個百分點排名第 02,Odyssey-3 則以 +9.99 個百分點排名第 03。因此,該廠商所謂的「最新技術水準」,在它自己引用的榜單上其實是第二名——這可能是因為該說法早於榜單更新,或是因為其涵蓋範圍比那句話所暗示的更狹隘。無論如何,誠實的解讀是:Odyssey-3 Pro 是排名前三的物理模型,而非絕對的領先者。FLUX 3 [large] 每支影片的成本也高得多,為 $0.868,而 Odyssey-3 Pro 為 $0.267;這正是該榜單的成本檢視所要揭露的取捨。

Odyssey 確實在子指標細分中拿下一個絕對第一:時空(Spatiotemporal),得分 44.70,領先 42.97 的 Odyssey-3 Pro 與 41.57 的 Physis-Lang(Cosmos3 Super)。在空間(Spatial)上它排名第四(59.17),落後 FLUX 3(64.36)、Odyssey-3 Pro(61.78)與 Physis-Lang(59.87);在加權空間(Weighted Spatial)與 MSE 上則分別排名第四與第三。這種模式——在動作隨時間維持一致性的表現上絕對領先,但在單幀空間保真度上屬於第二梯隊——是一個為了模擬演變而非繪製美麗靜態影像的模型所具備的一致特徵。

現在說說那個缺席的名字。Wan 3.0 完全沒有出現在 Physics-IQ Verified 上。榜上唯一的 Wan 條目是排名第 17 的 Wan 2.2 14B(−6.64 pp),以及排名第 22 的 Wan 2.2 5B(−11.13 pp)——兩者都低於該賽道的平均值,兩者都是開源,兩者都落後一個世代。該榜單自身的範圍說明指出,排行榜「涵蓋已受基準測試的模型,以及從預印本或發表公告中得知的模型,即使無法公開取用或尚未確認發布日期」,因此 Wan 3.0 的缺席並不是它得分很差的證據。這只證明沒有人在这个軸度上測量過它。實際後果很直白:任何聲稱 Odyssey-3 在物理合理性上勝過 Wan 3.0 的說法,在兩個方向上都是未經證實的;而任何聲稱它並未勝出的說法,同樣也未經證實。Wan 3.0 的獨立紀錄在別處——它在 OpenArt Arena 總排名拿下第 2,並在 Artificial Analysis 的 Video Editing with Audio 項目排名第 1——那些榜單衡量的是感知品質與剪輯品質,而非物理。

今天你可以稱呼的,以及你只能詢問的

Wan 3.0 has been callable since 24 August 2026, when A​libaba Cloud's launch turned the metered, application-gated beta into a fully open API. The model is reachable through the vendor's own API and several third-party platforms, all of them metered, with the per-second rates above. If a team needs generated video this afternoon, that is a real option and it comes with a bill.

Odyssey-3 並不提供那個。Odyssey 的頁面說研究預覽「現已可用」,並邀請實體 AI 開發者聯繫——這描述的是一場展示與對話,而不是一個藏在金鑰後方的端點。沒有公佈價格、沒有授權條款、沒有 API 文件,而且正如上一節所示,沒有獨立評估。開發者閱讀發佈文章時,今天無法計算在 Odyssey-3 上進行正式環境建置會花多少成本,也無法檢查那些物理學宣稱在供應商自家測試框架之外是否成立。這對一個發佈日的世界模型來說很正常,而且它也是這項比較中唯一最重要的事實。

由於這兩個模型都不在我們的目錄中——OrcaRouter 模型清單既不包含 Odyssey-3,也不包含 Wan 3.0——有用的路由切入點是位於它們之上的那一層。影片建置流程多半不會只呼叫一個模型。它會呼叫提示詞改寫模型、影片模型,有時還有音訊模型,而且每當供應商調整價格或速率限制時,就會重新呼叫所有這些模型。在 OrcaRouter 上,這些都位在 一個涵蓋 200 多個模型、以供應商定價提供且 0% 加價的 API之後,因此影片供應商按秒計價的降價當天就會在此生效,而不必等待設定變更,而且 自動容錯移轉表示某個模型中斷不會讓算圖佇列停滯。我們確實有提供的影片模型——MiniMax H3、Kling 3.0、Kling Video O1 及其他——都在同一組金鑰之下,這正是當 Wan 3.0 預覽版或 Odyssey-3 試用版終於開放時,替換成本之所以低廉的原因。

Capture of Odyssey's own 'Meet Odyssey-3' launch page, dated October 8th 2026, describing Odyssey-3 as a learned dynamical system implemented as an autoregressive diffusion transformer and inviting physical-AI developers to get in touch about the research preview.

哪一個屬於你的技術棧?

這個選擇並不接近,因為兩者所產出的東西毫無重疊。誠實的決策準則取決於你最終需要的產物。

如果交付成果是給人觀看的影片素材——廣告、插入鏡頭、社群短版、教學片段——就選擇 Wan 3.0。每次執行三十秒並具備原生音訊,跨圖像、影片與音訊的參考控制,以指令為基礎的編輯,以及可將文件作為輸入,這些構成了製作功能清單;而計量制的 API 意味著測試成本可事先得知。這是生成,不是模擬——當片段結束時,模型的工作也就結束。依廠商的每秒費率編列預算,並將廠商本身對音訊與畫面文字提出的注意事項視為設計限制,而非附註。

選擇 Odyssey-3——在你拿得到它的時候——如果交付成果是一種行為而非一個檔案:一個策略從中學習的環境、一個對動作的回應必須在長時間 rollout 中保持一致的場景、一項相機本身就是問題一部分的導航或操作任務。這正是時空合理性拿下絕對第一,以及每支影片 $0.139–$0.267 的模型化成本真正重要之處。缺少的是生產團隊做出承諾所需的一切——價格、授權、端點,以及外部量測——而這些都不是展示能替代的東西。

值得關注的重點很具體,而且近在眼前。如果 Wan 3.0 出現在 Physics-IQ Verified 上,物理問題就能朝一個方向獲得解答。如果 Odyssey 公佈費率表或 API,成本問題就會朝另一個方向解決。在這些事情之一發生之前,這項比較呈現出奇怪的樣貌:一個已公佈價格、卻沒有物理分數的正式生產系統,旁邊是一個已公佈物理分數、卻沒有價格的模擬器。任何人若告訴你今天哪個比較好,至少都是在猜這兩個數字中的其中一個。

Capture of the Physics-IQ Verified leaderboard from Anates Labs and DeepMind, ranked by net improvement in percentage points above the track mean, showing FLUX 3 [large] first at +12.27 pp, Odyssey-3 Pro second at +11.15 pp and Odyssey-3 third at +9.99 pp, with Wan 2.2 14B at rank 17 and Wan 2.2 5B at rank 22 and no entry for Wan 3.0.