
DeepSeek V4 Pro 對比 MiniMax M3:同樣的開放權重策略,價格卻高出五倍
- deepseek新DeepSeek: DeepSeek V4.1 Flash2026-09-1040智能
- openai新OpenAI: GPT-6 Astra2026-09-0453智能77程式
- google新Google: Gemini 3.8 Flash2026-09-0241智能76程式
- qwen新Qwen: Qwen3.8 Max (0902)2026-09-0240智能72程式
- anthropic新Anthropic: Claude Fable 5.12026-09-0153智能82程式
- AlibabaQwen: Qwen3.8 Flash2026-08-26$0.15 / $0.47 每百萬 tokens
- z-aiZ.ai: GLM 5.3 Flash2026-08-2642智能72程式
- DeepSeekDeepSeek: DeepSeek V4 Flash Vision (Exp)2026-08-21$0.22 / $0.66 每百萬 tokens
- z-aiZ.ai: GLM 5.32026-08-1845智能75程式
- obsidianQwen3.8 27B2026-08-1534智能68程式
- deepseekDeepSeek: DeepSeek V4 Pro 08132026-08-1236智能69程式
- grokSpaceXAI: Grok 4.62026-08-1244智能77程式
- metaMeta: Muse Spark 1.22026-08-0540智能72程式
- qwenQwen: Qwen3.8 Max2026-08-0340智能72程式
- deepseekDeepSeek: DeepSeek V4 Flash 07312026-07-3135智能69程式
- minimaxMiniMax: MiniMax-H32026-07-31minimax/minimax-h3
- qwenQwen: Qwen3.7 Flash2026-07-27$0.03 / $0.13 每百萬 tokens
- orcaOrcaDub: OrcaDub 1.02026-07-27orca/dub
- anthropicAnthropic: Claude Opus 52026-07-2451智能78程式
- googleGoogle: Gemini 3.6 Flash2026-07-2134智能69程式
DeepSeek V4 Pro and MiniMax M3 are both open-weights reasoning models with a million-token context, and they are both cheap enough that the word "budget" no longer means "weak." But their launches could not look more different: MiniMax M3 (MiniMax's 428B-parameter model, released June 2026) went out with a community license and a $0.30/$1.20 per million token price that undercut the entire field, while DeepSeek V4 Pro (released August 13, 2026, at $0.66/$1.98 off-peak, $1.32/$3.96 peak) arrived six weeks later with MIT-licensed weights and a slightly higher price. The twist is what happened to the more expensive one: between September 8 and 11, DeepSeek first scheduled V4 Pro for retirement, then delayed it, then cancelled the retirement entirely, committing to keep serving it at unchanged billing. So the matchup that a month ago looked like "cheap and cheaper" now has a durability story on the one side that no one expected — and that makes it worth re-reading the whole thing.
以下的獨立測量數據來自 Artificial Analysis。廠商回報的數字均已標示。
兩款平價開放權重旗艦,僅相差一個指數點
在 Artificial Analysis Intelligence Index 上,MiniMax M3 得分 30,在同級中排名第 16,而 DeepSeek V4 Pro 得分 36,排名第 7。兩款模型每 token 成本幾乎相同,卻有六分的差距。這就是重點:較貴的模型同時也是可測量上更聰明的那一個——但僅僅高出六分,而且只在 V4 Pro 所屬的純文字軸線上。MiniMax M3 可接受文字、圖像與影片輸入;V4 Pro 則僅支援文字。如果多模態輸入是必要條件,比較到此為止——兩者之中,只有 MiniMax M3 能讀取影片。
• 智慧指數 — V4 Pro 36(#7/113)對比 MiniMax M3 30(#16/113)
• 價格 — MiniMax M3 每 100 萬 $0.30/$1.20,對比 V4 Pro 離峰 $0.66/$1.98(尖峰 $1.32/$3.96)
• 速度 — MiniMax M3 約 103 詞元/秒,V4 Pro 約 81 詞元/秒
• 上下文 — 兩者皆為 1M tokens
• 參數 — MiniMax M3 總參數 428B / 活躍 23B,對比 V4 Pro 總參數 1.6T / 活躍 49B
• 輸入 — MiniMax M3 文字/圖片/影片 對比 V4 Pro 僅文字
• License — MiniMax Community License vs V4 Pro MIT
價格競爭比標價上看起來的更為接近。
直接捉對比較,MiniMax M3 的輸入價格比 V4 Pro 的離峰費率便宜 2.2 倍,輸出價格則便宜 1.65 倍。但牌價掩蓋了快取與尖峰的運作機制。V4 Pro 的離峰時段把原本就已低廉的價格再砍一半,而它的快取讀取折扣更高達約 97%——離峰時每百萬個快取 token 約 0.02 美元。MiniMax M3 的快取定價則沒有那麼積極。對於會重複使用大量前綴的工作負載——例如龐大的系統提示、長篇文件、檢索語料庫——V4 Pro 在離峰快取時段的實際每 token 成本,可能反而低於 MiniMax M3 的標示價格,這是對牌價排名的一大意外反轉。在全新 token 的工作負載上,MiniMax M3 較便宜;在快取比重高的工作負載上,V4 Pro 則可能勝出。
速度則呈現類似的反向情況。MiniMax M3 每秒約輸出 103 個 token,而 V4 Pro 為 81 個——是同價位層級中速度最快的模型,也比好幾款價格貴上五倍的模型更快。速度優勢確實存在,但並非極大;價格優勢更為重要,而且取決於工作負載。
授權與耐用性:九月改變答案之處
Both models are open weights, but the licenses differ. MiniMax M3 ships under the MINIMAX COMMUNITY LICENSE — permissive in spirit, but with its own terms. V4 Pro is MIT — about as unrestricted as an open-weights model gets, covering commercial use, fine-tuning, redistribution, and closed modifications. For enterprise adoption the MIT license is the cleaner story, and the gap shows up the moment your legal team reads the terms.
一個月前還是個真問題的持久性疑問,成了九月的頭條消息。DeepSeek 在九月上半段先是宣布 V4 Pro 將於 9 月 14 日退役,接著延後,然後在 11 日取消,並公開承諾以不變的計費方式繼續提供該模型服務,若有變動會提前通知。MiniMax M3 沒有這等戲劇性,因為本來就沒人預期會有:它是一款年輕的旗艦模型,短期內看不到退役的跡象。但效果是一樣的:這場對決的雙方如今都有了清楚的路線圖,而兩週前,其中一方還談不上有這樣的路線圖。

該由誰來選擇哪一個
如果你的工作負載包含圖像或影片輸入,答案已經很明確:MiniMax M3 是兩者中唯一能讀取這些內容的模型,而且價格是同級中最低的。如果你的工作負載是純文字且高度依賴快取——例如帶有大型系統提示的代理、長文件 RAG、大規模語料摘要——V4 Pro 的離峰快取時段與 MIT 授權,讓它成為每次回答成本效益更高的選擇,尤其它現在已確定會繼續留存。如果你的工作負載是純文字且對延遲敏感,MiniMax M3 的每秒 103 個 token 與 $0.30/$1.20 的價格,讓它成為目前市面上最快又便宜的選項。
兩個模型都能透過 OrcaRouter 以同一把金鑰取用,供應商定價以零加成原樣傳遞——所以上述數字就是你實際支付的數字——而供應商之間的自動容錯移轉,正是這場對決無法事先斷定的那件事的務實解答:哪個模型的供應商能在你的實際負載下撐得住。將多模態流量導向 MiniMax M3、把長脈絡文字工作導向 DeepSeek V4 Pro,只是兩行路由 DSL 的變更,而容錯移轉意味著任一方的供應商不穩都只是一次重試,而不是一起事故。


裁決
這是 DeepSeek V4 Pro 系列中最接近的一場對決,因為兩款模型都很便宜、都是開放權重,而且都很快。MiniMax M3 在原始價格、速度和多模態輸入方面勝出。DeepSeek V4 Pro 則在智慧、授權乾淨度,以及——自 9 月 11 日起——耐久性確定性方面勝出。一句話答案:多模態或延遲關鍵的文字交給 MiniMax M3;快取密集型、純文字或法律敏感的工作負載則交給 DeepSeek V4 Pro。
將多模態流量導向 MiniMax M3,並將長上下文文字工作導向 DeepSeek V4 Pro:MiniMax M3 只需修改兩行路由 DSL。
本文中的比較1
根據本文內容識別 · 基準測試:Artificial Analysis · 每日更新
