
DeepSeek V4 Pro לעומת Qwen3.8 Max: פער המחירים פי 3, הרצה חוזרת לאחר ההפוגה של ספטמבר
- deepseekחדשDeepSeek: DeepSeek V4.1 Flash2026-09-1040אינטליגנציה
- openaiחדשOpenAI: GPT-6 Astra2026-09-0453אינטליגנציה77כתיבת קוד
- googleחדשGoogle: Gemini 3.8 Flash2026-09-0241אינטליגנציה76כתיבת קוד
- qwenחדשQwen: Qwen3.8 Max (0902)2026-09-0240אינטליגנציה72כתיבת קוד
- anthropicחדשAnthropic: Claude Fable 5.12026-09-0153אינטליגנציה82כתיבת קוד
- AlibabaQwen: Qwen3.8 Flash2026-08-26$0.15 / $0.47 לכל 1M טוקנים
- z-aiZ.ai: GLM 5.3 Flash2026-08-2642אינטליגנציה72כתיבת קוד
- DeepSeekDeepSeek: DeepSeek V4 Flash Vision (Exp)2026-08-21$0.22 / $0.66 לכל 1M טוקנים
- z-aiZ.ai: GLM 5.32026-08-1845אינטליגנציה75כתיבת קוד
- obsidianQwen3.8 27B2026-08-1534אינטליגנציה68כתיבת קוד
- deepseekDeepSeek: DeepSeek V4 Pro 08132026-08-1236אינטליגנציה69כתיבת קוד
- grokSpaceXAI: Grok 4.62026-08-1244אינטליגנציה77כתיבת קוד
- metaMeta: Muse Spark 1.22026-08-0540אינטליגנציה72כתיבת קוד
- qwenQwen: Qwen3.8 Max2026-08-0340אינטליגנציה72כתיבת קוד
- deepseekDeepSeek: DeepSeek V4 Flash 07312026-07-3135אינטליגנציה69כתיבת קוד
- minimaxMiniMax: MiniMax-H32026-07-31minimax/minimax-h3
- qwenQwen: Qwen3.7 Flash2026-07-27$0.03 / $0.13 לכל 1M טוקנים
- orcaOrcaDub: OrcaDub 1.02026-07-27orca/dub
- anthropicAnthropic: Claude Opus 52026-07-2451אינטליגנציה78כתיבת קוד
- googleGoogle: Gemini 3.6 Flash2026-07-2134אינטליגנציה69כתיבת קוד
DeepSeek V4 Pro and Qwen3.8 Max are the two most important open-or-cheap flagships of this generation, and the September events on the DeepSeek side changed the matchup more than either vendor's marketing has acknowledged. DeepSeek V4 Pro shipped on August 13, 2026, ten days after Alibaba's Qwen3.8 Max went generally available on August 3, and the pairing has always read as a one-sided contest: Qwen3.8 Max is the 2.4-trillion-parameter all-rounder at $2 per million input and $6 per million output tokens, while DeepSeek V4 Pro is the 1.6-trillion-parameter reasoning specialist at a fraction of that price, MIT-licensed, with public weights. The reason this comparison is worth re-running today is not the launch-week numbers but what happened a month later — DeepSeek spent the week of September 8 trying to retire V4 Pro, then reversed on September 11 and committed to keeping it online at unchanged billing. That reversal is the single most important fact about this matchup right now, because it decides whether the cheap half is a safe bet or a migration trap.
Independent measurements in this piece come from Artificial Analysis, captured today. Alibaba's and DeepSeek's own benchmark tables are labeled vendor-reported throughout — neither company's launch numbers have been fully reproduced by a third party, and DeepSeek's are now four weeks old against a moving scoreboard.
ההפוגה בספטמבר היא הסיפור, לא ההשקה באוגוסט
ב-8 בספטמבר הודיעה DeepSeek שאחרי 14 בספטמבר כל בקשה אל deepseek-v4-pro תנותב אל V4.1 Flash ותחויב במחירי Flash עד שיצא V4.1 Pro. כל מי שהריץ השוואה רצינית בין Qwen3.8 Max ל-DeepSeek V4 Pro היה צריך להתחיל לתכנן לקראת היעלמותה של האפשרות הזולה. ואז DeepSeek דחתה את מועד הניתוק, וב-11 בספטמבר היא חזרה בה לחלוטין, והודיעה שתמשיך לספק שירותי API עבור DeepSeek V4 Pro גם אחרי 14 בספטמבר, עם חיוב ללא שינוי, ותיתן הודעה אם זה ישתנה אי פעם. דבר המהלך ההפוך דווח באותו יום על ידי כלי התקשורת הממלכתיים בסין והעיתונות הטכנולוגית.
מה שזה עושה להשוואה הוא להסיר את האסימטריה שכיוונה בשקט את ההחלטות. במהלך השבוע הראשון של ספטמבר, התשובה הכנה לשאלה "האם לבנות על מודל הפלט ב-$0.66 או על מודל הפלט ב-$6?" זוהם בסיכון הישרדות: למוביל המחיר היה תאריך כיבוי. נכון ל-15 בספטמבר, הזיהום הזה נעלם. האפשרות הזולה עמידה, וההשוואה סוף סוף עוסקת במה שהיא תמיד הייתה אמורה לעסוק בו — אינטליגנציה, מחיר ומהירות.

מחיר: אותו משחק, בשבריר מעלות הפלט
בלוח התוצאות העצמאי הנוכחי השניים מרוחקים בארבע נקודות, ופער המחירים ביניהם הוא הגדול ביותר בדרג המשקלים הפתוחים. Qwen3.8 Max עולה $2.00 למיליון אסימוני קלט ו-$6.00 למיליון אסימוני פלט במחיר המחירון של Alibaba, עם הנחת מטמון של 88%. DeepSeek V4 Pro עולה $0.66 למיליון אסימוני קלט ו-$1.98 למיליון אסימוני פלט בשעות שפל, ומכפיל את עצמו ל-$1.32/$3.96 בשעות שיא, עם הנחת מטמון של 97%.
• מחיר — DeepSeek V4 Pro $0.66/$1.98 לכל 1M בשעות שפל (בשעות שיא $1.32/$3.96) לעומת Qwen3.8 Max $2.00/$6.00 אחיד
• מדד אינטליגנציה — V4 Pro 36 (#7/113) לעומת Qwen3.8 Max 40 (#30/200)
• מהירות פלט — V4 Pro ~81 טוקנים/שנייה לעומת Qwen3.8 Max ~41 טוקנים/שנייה
• עלות לכל משימה — V4 Pro $0.67 לעומת Qwen3.8 Max $2.67 (לכל משימת אינדקס AA)
• הקשר — שניהם 1M טוקנים
• רישיון — V4 Pro עם משקולות פתוחות MIT לעומת Qwen3.8 Max עם משקולות קנייניות, רישיון מסחרי
The output-price ratio is the whole argument: Qwen3.8 Max costs three times as much per input token and roughly three times per output token at V4 Pro's peak rate — and the gap on output widens to six times when V4 Pro is billed at its off-peak rate, which covers most of the week outside a narrow window. On a reasoning-heavy workload, where output tokens dominate the bill, that gap is the difference between a model you can leave running and one you watch in the dashboard. The cache lines widen it further: V4 Pro's 97% discount against Qwen's 88% means a long-context task with a reused prefix costs a fraction of its list price on the DeepSeek side.
אינטליגנציה: ארבע נקודות פער, והפער הוא המוצר האמיתי
The independent scoreboard makes this closer than the price ratio suggests. On the current Artificial Analysis Intelligence Index, Qwen3.8 Max scores 40 against DeepSeek V4 Pro's 36 — a four-point gap that shows up in reasoning-heavy work but disappears in token-heavy work. Qwen3.8 Max's edge is real and consistent: it ranks #30 of 200 overall, just inside the frontier, while V4 Pro ranks #7 of 113 in its open-weights class. On Alibaba's own benchmark table, Qwen3.8 Max claims a Terminal-Bench 2.1 of 86.6, GPQA Diamond 92.6, and a FrontierSWE 73.5 that nearly doubled its predecessor's 40.7 — all vendor-reported, none reproduced independently. DeepSeek's model card claims a Codeforces rating of 3,348 and a Terminal-Bench 2.1 of 87.9 at maximum reasoning effort — likewise vendor-reported and unverified by a third party.
The honest framing is that the two companies are arguing past each other. Qwen3.8 Max's vendor table is built around enterprise and scientific work — the 0902 refresh Alibaba shipped on September 2 was further post-trained on Coding and Cowork, which Qwen says strengthens exactly that profile. DeepSeek V4 Pro's own claims center on deep reasoning and agentic tool use, where its four-week-old Codeforces and Terminal-Bench numbers still look strong. Neither vendor's table has been reproduced by an independent evaluator, and on the one scoreboard that is independent, the four-point gap is the whole difference between the models.
המהירות היא המקום שבו Qwen3.8 Max מאבד קרקע בשקט
הממד שדף המפרט מסתיר הוא מהירות הפלט. Artificial Analysis מודדת את DeepSeek V4 Pro בכ-81 טוקנים לשנייה לעומת כ-41 של Qwen3.8 Max — פער של פי שניים במהירות שבה מתקבלות התשובות. בעומס שיחה זהו הפסקה שהמשתמש מבחין בה; בלולאת סוכן עם קריאות עוקבות רבות, זה מצטבר לזמן אמיתי בשעון. יחס המחירים כבר מעדיף את V4 Pro, ופער המהירות הופך את פער העלות האפקטיבית לתשובה לרחב עוד יותר מכפי שמתמטיקת הטוקנים מרמזת.
יש פשרה טמונה במספר הזה, והיא הפַּטְפְּטָנוּת של Qwen3.8 Max. Artificial Analysis מציינת שהמודל "איטי באופן ניכר ומאוד מפורט", והעלות שלו לכל משימה — 2.67 דולר לעומת 0.67 דולר של V4 Pro — משקפת גם את התעריף הגבוה יותר וגם את טוקני הפלט הנוספים. מודל חשיבה שכותב יותר בכל תשובה עולה יותר לכל משימה עוד לפני שמכפילים במחירון. אם עומס העבודה שלך רגיש לזמן תגובה, או שלולאת הסוכן שלך מבצעת עשרות קריאות, זה המספר שעליו כדאי להריץ הערכה משלך.
משקלים: MIT לעומת רישיון מסחרי
DeepSeek V4 Pro is MIT-licensed with public weights — you can download them, serve them yourself, fine-tune them, and keep your modifications closed. Qwen3.8 Max is the first Max-class Qwen ever to ship open weights (the 2.4T A95B checkpoint landed on August 12), but under a commercial license with scale-tier conditions, and Alibaba has not published the full inference stack the way DeepSeek has. For a company comparing the two, the deployment story is the tie-breaker: V4 Pro can be self-hosted as infrastructure; Qwen3.8 Max is primarily a service you rent. The weights difference matters most if your concern is vendor lock-in or if you want to escape the API price entirely.
מי צריך לבחור איזה
אם עומס העבודה שלך עתיר הסקה, נשלט על ידי טוקנים של פלט, או רגיש להשהיה, DeepSeek V4 Pro הוא התשובה כמעט בכל ציר שמופיע בחשבון: זול בערך פי שלושה עד שישה בפלט, מהיר בערך פי שניים, הנחת מטמון עמוקה יותר, ומשקלות MIT כדרך מילוט. פער הבינה של ארבע נקודות הוא אמיתי אך צר, והוא לעיתים רחוקות שורד מפגש עם פער המחירים בתעבורת ייצור.
אם עומס העבודה שלך הוא ארגוני באופיו — ניתוח מורכב רב-שלבי, הסקה לוגית עמוקה, עבודה אג'נטית תובענית, או כל משימה שבה איכות התשובה השולית שווה פי שלושה עד שישה מעלות הטוקנים — ארבע הנקודות הנוספות של Qwen3.8 Max במדד העצמאי והטענות החזקות יותר שלו בנוגע לביצועים ארגוניים הם הסיבה לשלם את הפרמיה. רענון 0902 מחדד את הפרופיל הזה. ואם אתה זקוק לקלט של תמונה או וידאו, הבחירה בכלל לא קרובה: Qwen3.8 Max מקבל טקסט, תמונה ווידאו, בעוד DeepSeek V4 Pro הוא טקסט בלבד.
שני המודלים ניתנים לקריאה דרך OrcaRouter באמצעות מפתח אחד — גם DeepSeek V4 Pro וגם Qwen3.8 Max מנותבים דרך הפלטפורמה — ומכיוון ש-OrcaRouter מעביר את מחירי המחירון של הספקים ללא תוספת, המספרים במאמר זה הם המספרים שאתם משלמים בפועל, כאשר הורדות המחירים של הספקים נכנסות לתוקף באותו היום. ה-DSL לניתוב מאפשר לכם לשלוח את התעבורה עתירת הטוקנים או רגישת ההשהיה אל DeepSeek V4 Pro, ואת המשימות הרב-מודאליות או בעלות הסיכון הגבוה ביותר אל Qwen3.8 Max, מתוך אותה אינטגרציה — וכך בדיוק רוב הצוותים מגיעים בסופו של דבר להשתמש בצירוף הזה: את שניהם, כל אחד למה שהוא עושה הכי טוב.


פסק הדין
ההפוגה בספטמבר הכריעה את השאלה שכבר הטביעה את ההתמודדות הזו: החצי הזול אינו עוד סיכון נטישה. DeepSeek V4 Pro הוא התשובה של עלות-ביצועים — זול פי שלושה עד שישה בתפוקה, מהיר פי שניים, ברישיון MIT, וכעת מחויב להישאר זמין. Qwen3.8 Max הוא תשובת היכולות — גבוה בארבע נקודות מדד עצמאיות, טענות ארגוניות חזקות יותר, קלט רב-מודאלי, ורישיון מסחרי שמונע ממנו להיכנס לסטאק שלך. בחר בארבע הנקודות כשההערכות שלך מוכיחות שהן חשובות; בחר במחיר כשהחשבון שלך הוא ההערכה שבאמת מכריעה.
שלח את התעבורה עתירת הטוקנים או הרגישה להשהיה אל DeepSeek V4 Pro ואת הסקה המולטימודלית או בעלת הסיכון הגבוה ביותר אל Qwen3.8 Max מתוך אותה אינטגרציה — וזה בדיוק האופן שבו רוב הצוותים משתמשים בסופו של דבר בצירוף הזה: שניהם, כל אחד למה שהוא עושה הכי טוב.
השוואות במאמר הזה1
זוהה מתוך המאמר הזה · בנצ'מרקים: Artificial Analysis · מתעדכן יומית
