
Vidu Q4プレビュー対Alibaba Wan 2.7:2つのモード対4つ、そして{{1}}本当に重要な{{/1}}{{2}}場面{{/2}}での一つの評価
- openaiNEWOpenAI: GPT-6.1 Sol2026-09-2952知能
- anthropicNEWAnthropic: Claude Sonnet 5.52026-09-2856知能
- typesafeNEWTypeSafe: Jev 1.132026-09-24$0.04 / $0.00 100万トークンあたり · 128 tok/s
- OpenAIOpenAI: GPT-6 Luna2026-09-2238知能
- OpenAIOpenAI: GPT-6 Sol2026-09-2248知能
- AnthropicAnthropic: Claude Opus 5.52026-09-2258知能
- xAIGrok 4.72026-09-2146知能
- OrcaOrca: OrcaCyber Zero 1.02026-09-17$3.00 / $7.50 100万トークンあたり · 62 tok/s
- OrcaOrca: OrcaVerify Text 1.02026-09-16$2.00 / $0.00 100万トークンあたり · 320 tok/s
- DeepSeekDeepSeek: DeepSeek V4.1 Flash2026-09-1040知能
- OpenAIOpenAI: GPT-6 Astra2026-09-0453知能77コーディング
- GoogleGoogle: Gemini 3.8 Flash2026-09-0241知能76コーディング
- AlibabaQwen: Qwen3.8 Max (0902)2026-09-0245知能76コーディング
- AnthropicAnthropic: Claude Fable 5.12026-09-0153知能82コーディング
- TencentTencent: Hy4 preview2026-08-28$0.83 / $2.50 100万トークンあたり · 54 tok/s
- AlibabaQwen: Qwen3.8 Flash2026-08-26$0.15 / $0.47 100万トークンあたり · 358 tok/s
- z-aiZ.ai: GLM 5.3 Flash2026-08-2642知能72コーディング
- DeepSeekDeepSeek: DeepSeek V4 Flash Vision (Exp)2026-08-21$0.22 / $0.66 100万トークンあたり · 232 tok/s
- z-aiZ.ai: GLM 5.32026-08-1845知能75コーディング
- obsidianQwen3.8 27B2026-08-1534知能68コーディング
Here is the whole comparison in one line: Vidu Q4 Preview beats Alibaba Wan 2.7 by 102 Elo on the Artificial Analysis image-to-video board, and cannot be entered against it on text-to-video or video editing, because Vidu Q4 Preview does not do either.
Both models launched in 2026 and both generate video with sound. Vidu Q4 Preview arrived on 7 October 2026 from Shengshu Technology as a first public preview ahead of the full Q4 model, in two modes — Image-to-Video and Reference-to-Video. Alibaba Wan 2.7 shipped in April 2026 as a four-variant suite: text-to-video, image-to-video, reference-to-video and video edit, billed per second of generated video. So the tie-breaker in most head-to-head pages — a score — only exists on one of the axes, and the honest way to judge these two is to start by asking which axes you actually need.
隙間が見えたままのスコアボード
Artificial Analysisは3つの独立した動画リーダーボードを運営している。それぞれが独自のEloスケールと独自の投票プールを備えているため、あるボードのスコアが別のボードに引き継がれることはない。2026年10月8日閲覧、今回も音声を保持した状態で:
• Image-to-video (AA-Video-I2V v1.0) — Vidu Q4 Preview 3rd, 1,179 ±10, 5,543 samples, Oct 2026, $7.20/min; Alibaba Wan 2.7 13th, 1,077 ±8, 4,571 samples, Apr 2026, $9.00/min
• テキストから動画(AA-Video-T2V v2.0) — Vidu Q4 Preview は未収録、Wan2.7-260612 は13~14位、1,030 ±9、5,571サンプル、2026年6月、$9.00/分
• 動画編集(AA-Video-Editing v1.1) — Vidu Q4 Preview は未収録、Wan 2.7 は7位、1,074、20,021サンプル、2026年4月、$16.90/分
その表から2つのことが導かれる。Wan 2.7は2つの中で圧倒的に汎用性が高く、3つのボードすべてでスコアが付けられている唯一のモデルであり、その編集部門のエントリーは、どこにあるWan 2.7のスコアよりも最多の投票数を集めており、2万サンプルに及ぶ。そして両者が交わる唯一の軸では、Vidu Q4 Previewが102 Elo差で勝利しており、これはどちらかのモデルの信頼区間の幅のおよそ5倍にあたる。
The 102-point gap is not a small result. For scale: the whole spread across the top six places on the image-to-video board is 21 Elo. Vidu Q4 Preview is not marginally ahead of Wan 2.7 on image-to-video; it is a different tier, and it costs $1.80 per minute less while also being the more recent release. If image-to-video is your workload, the decision is not close and the independent board is the reason it is not close.

2つの中でWan 2.7だけが表示される場合
Wan 2.7 の自社ドキュメントは、まだ改良が必要な領域として音声品質と画面上のテキストの正確さの2つを挙げている。これはレビュアーではなくベンダー自身による評価であり、どの比較においても心に留めておく価値がある。というのも、両モデルとも「後から後付けするのではなく、映像とともに生成されるネイティブ音声」という同じ看板の主張を掲げているからだ。
The difference is in what happens by default. Vidu Q4 Preview's image-to-video endpoint carries an audio parameter that defaults to true: you get sound unless you turn it off, and turning it off is how you get a silent clip. Wan 2.7 does not document the same switch in the same place. If silent output matters to your pipeline — and for anything you plan to score separately, or dub into another language, it does — that asymmetry changes what your first integration pass looks like.
編集ボードでは、20,021サンプル中1,074票で7位というWan 2.7の結果は、投票数で見た場合、どちらのモデルであれどのボードでも収めた中で最も強い単一の結果です。2万票は、背後に確かな重みのある測定であり、そのボードには比較対象となるViduモデルは現れていません。あなたの作業が新しいショットを生成するのではなく、既存の映像のタイミングを変えたりスタイルを変えたりするものであるなら、Wan 2.7は2つのうち唯一該当するものです。

生成側の仕様ギャップ
重なり合う部分において、両者はほとんどあらゆる次元で異なっており、その違いはWanに有利な形で対称的ではない。
• 最大解像度 — Vidu Q4 Preview 4K(2Kおよび4K、10ビットカラー)対 Wan 2.7 は最大1080p
• 最大クリップ長 — Vidu Q4 Preview 16秒(画像から動画 3~16秒、参照から動画 1~16秒) vs Wan 2.7 最大15秒
• 参照画像 — Vidu Q4 Preview は1回の設定で最大15枚、Wan 2.7 は最大10アセット
• Reference audio — Vidu Q4 Preview up to 3 clips for voice consistency vs not published to the same detail by Alibaba
• タスク対応範囲 — Vidu Q4 Preview の2モード(画像から動画、参照から動画)対 Wan 2.7 の4バリアント(テキストから動画と動画編集を追加)
• 公表リストレート — Vidu Q4 Preview はプロモーション価格で $0.014/秒から。独立系ボードでの実測は $7.20/分。対する Wan 2.7 はバリアント別に $0.10~$0.15/秒、実測 $9.00/分
解像度の項目は、慎重に比較検討すべき点です。10ビットでの4K出力は、スクリーンや大型ディスプレイ向けに仕上げる人にとって、成果物に真の違いをもたらすものであり、Wan 2.7がいずれのバリアントでも提供する機能ではありません。しかし4Kは、秒単位の価格設定が通常宣伝されている最低価格のようには見えなくなる領域でもあり、Vidu自身の発表資料には、最終価格、対応解像度、機能の利用可否、利用条件はプランと地域によって異なるという注意書きが付いています。

どちらですか、ワークロード別で
静止画をアニメーション化するなら — 商品写真、キャラクターの参照フレーム、絵コンテ — Vidu Q4 Preview は、入手可能な唯一の独立した証拠において、より優れたモデルであり、しかも測定された料金はより低い。それが明快な結論だ。
テキストから動画、画像から動画、そして編集をカバーする1つの契約が必要なら、Wan 2.7は2つのうち唯一それに応えるモデルであり、しかもどちらのモデルが得た投票数よりも多い投票数に基づく編集部門7位の結果で応えています。1つのベンダーに統合することは確かな利点であり、この比較はそうでないふりをしていません。
If you need both, the cost of running two vendors is usually not the licence — it is the second integration, the second key, the second set of retry semantics, and the second schema to keep current. OrcaRouter collapses that: one OpenAI-compatible endpoint across 200-plus models, provider list prices passed through with no markup added, and automatic failover across providers. On the video side specifically we serve MiniMax H3, which is the model currently ranked first and second on the image-to-video board above — callable on the same endpoint and the same key as the text models. Neither Vidu Q4 Preview nor Alibaba Wan 2.7 is an OrcaRouter route today, and this page should not be read as implying otherwise; what we can do is keep the rest of the workload on one integration so the video experiment is a line item rather than a project.
何がこれを変えるでしょうか
Vidu Q4 Previewはプレビュー版です。価格設定はプロモーション用で、ベンダー自身の説明によればこのビルドはQ4正式リリースより先行しており、APIドキュメントには正しいモデルエイリアスと並んで誤入力されたモデルエイリアスが依然として残っています。重要なのは2つの結果です。Q4正式リリースでテキストから動画への生成が追加されれば、評価軸が収束し、これは1つではなく3つのボードでの直接比較になります。そうでなければ、この分岐はそのまま残り、選択は純粋にワークロードの形態次第です。
On the other side, Wan 2.7's samples on the image-to-video board are five months older than Vidu's and about a thousand fewer. Both intervals will narrow. Neither model's rank should be treated as fixed beyond the date printed beside it.
短い答え:両者が共有する競争軸はただ一つだけで、そこでは Vidu Q4 Preview がより低い計測レートで明確にリードしており、Vidu が競合していないあらゆる領域では Wan 2.7 が事実上無条件で勝つ。長い答えはこうだ。2モードのプレビューモデルと4バリアントのスイートは、本来互いの代替ではない。それぞれ別の仕事に対する代替であり、この比較で役に立つのは、自分がどちらの仕事を抱えているかを知ることだ。
