כרטיס גיבור מחולל שכותרתו 'DeepSeek V4 Pro vs Qwen3.8 Max' עם שני כרטיסים מעוגלים — כרטיס DeepSeek V4 Pro שעליו תג '$0.66 / $1.98 בשעות שפל · משקלים פתוחים של MIT · הקשר של 1M' וכרטיס Qwen3.8 Max שעליו תג '$2.00 / $6.00 · רישיון מסחרי · הקשר של 1M' — מופרדים על ידי סמל מאזניים ובפוטר 'פי שלושה מהמחיר, ארבע נקודות של תבונה — מה שורד מפגש עם החשבון שלך?'
Guides & Insights

DeepSeek V4 Pro לעומת Qwen3.8 Max: פער המחירים פי 3, הרצה חוזרת לאחר ההפוגה של ספטמבר

מחבר

Rowan Sterling

תאריך פרסום

מודלים אחרונים · 20צפו בכל המודלים
בנצ'מרקים: Artificial Analysis · מתעדכן יומית
חזרה לכל הפוסטים

DeepSeek V4 Pro and Qwen3.8 Max are the two most important open-or-cheap flagships of this generation, and the September events on the Deep​Seek side changed the matchup more than either vendor's marketing has acknowledged. DeepSeek V4 Pro shipped on August 13, 2026, ten days after Aliba​ba's Qwen3.8 Max went generally available on August 3, and the pairing has always read as a one-sided contest: Qwen3.8 Max is the 2.4-trillion-parameter all-rounder at $2 per million input and $6 per million output tokens, while DeepSeek V4 Pro is the 1.6-trillion-parameter reasoning specialist at a fraction of that price, MIT-licensed, with public weights. The reason this comparison is worth re-running today is not the launch-week numbers but what happened a month later — Deep​Seek spent the week of September 8 trying to retire V4 Pro, then reversed on September 11 and committed to keeping it online at unchanged billing. That reversal is the single most important fact about this matchup right now, because it decides whether the cheap half is a safe bet or a migration trap.

Independent measurements in this piece come from Artificial Analysis, captured today. Aliba​ba's and Deep​Seek's own benchmark tables are labeled vendor-reported throughout — neither company's launch numbers have been fully reproduced by a third party, and Deep​Seek's are now four weeks old against a moving scoreboard.

ההפוגה בספטמבר היא הסיפור, לא ההשקה באוגוסט

ב-8 בספטמבר הודיעה DeepSeek שאחרי 14 בספטמבר כל בקשה אל deepseek-v4-pro תנותב אל V4.1 Flash ותחויב במחירי Flash עד שיצא V4.1 Pro. כל מי שהריץ השוואה רצינית בין Qwen3.8 Max ל-DeepSeek V4 Pro היה צריך להתחיל לתכנן לקראת היעלמותה של האפשרות הזולה. ואז DeepSeek דחתה את מועד הניתוק, וב-11 בספטמבר היא חזרה בה לחלוטין, והודיעה שתמשיך לספק שירותי API עבור DeepSeek V4 Pro גם אחרי 14 בספטמבר, עם חיוב ללא שינוי, ותיתן הודעה אם זה ישתנה אי פעם. דבר המהלך ההפוך דווח באותו יום על ידי כלי התקשורת הממלכתיים בסין והעיתונות הטכנולוגית.

מה שזה עושה להשוואה הוא להסיר את האסימטריה שכיוונה בשקט את ההחלטות. במהלך השבוע הראשון של ספטמבר, התשובה הכנה לשאלה "האם לבנות על מודל הפלט ב-$0.66 או על מודל הפלט ב-$6?" זוהם בסיכון הישרדות: למוביל המחיר היה תאריך כיבוי. נכון ל-15 בספטמבר, הזיהום הזה נעלם. האפשרות הזולה עמידה, וההשוואה סוף סוף עוסקת במה שהיא תמיד הייתה אמורה לעסוק בו — אינטליגנציה, מחיר ומהירות.

A generated scoreboard titled 'DeepSeek V4 Pro vs Qwen3.8 Max — the scoreboard': left column DeepSeek V4 Pro rows AA Index 36 (#7/113), Price $0.66/$1.98 off-peak, Speed ~81 tok/s, Context 1M, License MIT open weights, Input text; right column Qwen3.8 Max rows AA Index 40 (#30/200), Price $2.00/$6.00, Speed ~41 tok/s, Context 1M, License Commercial, Input text + image + video; footer 'AA figures independent (current scale); DeepSeek price is vendor list off-peak.'

מחיר: אותו משחק, בשבריר מעלות הפלט

בלוח התוצאות העצמאי הנוכחי השניים מרוחקים בארבע נקודות, ופער המחירים ביניהם הוא הגדול ביותר בדרג המשקלים הפתוחים. Qwen3.8 Max עולה $2.00 למיליון אסימוני קלט ו-$6.00 למיליון אסימוני פלט במחיר המחירון של Aliba​ba, עם הנחת מטמון של 88%. DeepSeek V4 Pro עולה $0.66 למיליון אסימוני קלט ו-$1.98 למיליון אסימוני פלט בשעות שפל, ומכפיל את עצמו ל-$1.32/$3.96 בשעות שיא, עם הנחת מטמון של 97%.

• מחיר — DeepSeek V4 Pro $0.66/$1.98 לכל 1M בשעות שפל (בשעות שיא $1.32/$3.96) לעומת Qwen3.8 Max $2.00/$6.00 אחיד

• מדד אינטליגנציה — V4 Pro 36 (#7/113) לעומת Qwen3.8 Max 40 (#30/200)

• מהירות פלט — V4 Pro ~81 טוקנים/שנייה לעומת Qwen3.8 Max ~41 טוקנים/שנייה

• עלות לכל משימה — V4 Pro $0.67 לעומת Qwen3.8 Max $2.67 (לכל משימת אינדקס AA)

• הקשר — שניהם 1M טוקנים

• רישיון — V4 Pro עם משקולות פתוחות MIT לעומת Qwen3.8 Max עם משקולות קנייניות, רישיון מסחרי

The output-price ratio is the whole argument: Qwen3.8 Max costs three times as much per input token and roughly three times per output token at V4 Pro's peak rate — and the gap on output widens to six times when V4 Pro is billed at its off-peak rate, which covers most of the week outside a narrow window. On a reasoning-heavy workload, where output tokens dominate the bill, that gap is the difference between a model you can leave running and one you watch in the dashboard. The cache lines widen it further: V4 Pro's 97% discount against Qw​en's 88% means a long-context task with a reused prefix costs a fraction of its list price on the Deep​Seek side.

אינטליגנציה: ארבע נקודות פער, והפער הוא המוצר האמיתי

The independent scoreboard makes this closer than the price ratio suggests. On the current Artificial Analysis Intelligence Index, Qwen3.8 Max scores 40 against DeepSeek V4 Pro's 36 — a four-point gap that shows up in reasoning-heavy work but disappears in token-heavy work. Qwen3.8 Max's edge is real and consistent: it ranks #30 of 200 overall, just inside the frontier, while V4 Pro ranks #7 of 113 in its open-weights class. On Aliba​ba's own benchmark table, Qwen3.8 Max claims a Terminal-Bench 2.1 of 86.6, GPQA Diamond 92.6, and a FrontierSWE 73.5 that nearly doubled its predecessor's 40.7 — all vendor-reported, none reproduced independently. Deep​Seek's model card claims a Codeforces rating of 3,348 and a Terminal-Bench 2.1 of 87.9 at maximum reasoning effort — likewise vendor-reported and unverified by a third party.

The honest framing is that the two companies are arguing past each other. Qwen3.8 Max's vendor table is built around enterprise and scientific work — the 0902 refresh Aliba​ba shipped on September 2 was further post-trained on Coding and Cowork, which Qw​en says strengthens exactly that profile. DeepSeek V4 Pro's own claims center on deep reasoning and agentic tool use, where its four-week-old Codeforces and Terminal-Bench numbers still look strong. Neither vendor's table has been reproduced by an independent evaluator, and on the one scoreboard that is independent, the four-point gap is the whole difference between the models.

המהירות היא המקום שבו Qwen3.8 Max מאבד קרקע בשקט

הממד שדף המפרט מסתיר הוא מהירות הפלט. Artificial Analysis מודדת את DeepSeek V4 Pro בכ-81 טוקנים לשנייה לעומת כ-41 של Qwen3.8 Max — פער של פי שניים במהירות שבה מתקבלות התשובות. בעומס שיחה זהו הפסקה שהמשתמש מבחין בה; בלולאת סוכן עם קריאות עוקבות רבות, זה מצטבר לזמן אמיתי בשעון. יחס המחירים כבר מעדיף את V4 Pro, ופער המהירות הופך את פער העלות האפקטיבית לתשובה לרחב עוד יותר מכפי שמתמטיקת הטוקנים מרמזת.

יש פשרה טמונה במספר הזה, והיא הפַּטְפְּטָנוּת של Qwen3.8 Max. Artificial Analysis מציינת שהמודל "איטי באופן ניכר ומאוד מפורט", והעלות שלו לכל משימה — 2.67 דולר לעומת 0.67 דולר של V4 Pro — משקפת גם את התעריף הגבוה יותר וגם את טוקני הפלט הנוספים. מודל חשיבה שכותב יותר בכל תשובה עולה יותר לכל משימה עוד לפני שמכפילים במחירון. אם עומס העבודה שלך רגיש לזמן תגובה, או שלולאת הסוכן שלך מבצעת עשרות קריאות, זה המספר שעליו כדאי להריץ הערכה משלך.

משקלים: MIT לעומת רישיון מסחרי

DeepSeek V4 Pro is MIT-licensed with public weights — you can download them, serve them yourself, fine-tune them, and keep your modifications closed. Qwen3.8 Max is the first Max-class Qw​en ever to ship open weights (the 2.4T A95B checkpoint landed on August 12), but under a commercial license with scale-tier conditions, and Aliba​ba has not published the full inference stack the way Deep​Seek has. For a company comparing the two, the deployment story is the tie-breaker: V4 Pro can be self-hosted as infrastructure; Qwen3.8 Max is primarily a service you rent. The weights difference matters most if your concern is vendor lock-in or if you want to escape the API price entirely.

מי צריך לבחור איזה

אם עומס העבודה שלך עתיר הסקה, נשלט על ידי טוקנים של פלט, או רגיש להשהיה, DeepSeek V4 Pro הוא התשובה כמעט בכל ציר שמופיע בחשבון: זול בערך פי שלושה עד שישה בפלט, מהיר בערך פי שניים, הנחת מטמון עמוקה יותר, ומשקלות MIT כדרך מילוט. פער הבינה של ארבע נקודות הוא אמיתי אך צר, והוא לעיתים רחוקות שורד מפגש עם פער המחירים בתעבורת ייצור.

אם עומס העבודה שלך הוא ארגוני באופיו — ניתוח מורכב רב-שלבי, הסקה לוגית עמוקה, עבודה אג'נטית תובענית, או כל משימה שבה איכות התשובה השולית שווה פי שלושה עד שישה מעלות הטוקנים — ארבע הנקודות הנוספות של Qwen3.8 Max במדד העצמאי והטענות החזקות יותר שלו בנוגע לביצועים ארגוניים הם הסיבה לשלם את הפרמיה. רענון 0902 מחדד את הפרופיל הזה. ואם אתה זקוק לקלט של תמונה או וידאו, הבחירה בכלל לא קרובה: Qwen3.8 Max מקבל טקסט, תמונה ווידאו, בעוד DeepSeek V4 Pro הוא טקסט בלבד.

שני המודלים ניתנים לקריאה דרך OrcaRouter באמצעות מפתח אחד — גם DeepSeek V4 Pro וגם Qwen3.8 Max מנותבים דרך הפלטפורמה — ומכיוון ש-OrcaRouter מעביר את מחירי המחירון של הספקים ללא תוספת, המספרים במאמר זה הם המספרים שאתם משלמים בפועל, כאשר הורדות המחירים של הספקים נכנסות לתוקף באותו היום. ה-DSL לניתוב מאפשר לכם לשלוח את התעבורה עתירת הטוקנים או רגישת ההשהיה אל DeepSeek V4 Pro, ואת המשימות הרב-מודאליות או בעלות הסיכון הגבוה ביותר אל Qwen3.8 Max, מתוך אותה אינטגרציה — וכך בדיוק רוב הצוותים מגיעים בסופו של דבר להשתמש בצירוף הזה: את שניהם, כל אחד למה שהוא עושה הכי טוב.

A screenshot of the Artificial Analysis page for Qwen3.8 Max, captured September 15, 2026, showing the Intelligence Index score of 40 at rank 30 of 200, output speed 40.6 tokens per second, an input price of $2.00 and output price of $6.00 per million tokens with an 88% cache discount, and a cost of $2.67 per Intelligence Index task.A screenshot of the OrcaRouter model page for Qwen3.8 Max at orcarouter.ai/models/qwen/qwen3.8-max, captured September 15, 2026, showing the model listing with its provider Qwen, 1M-token context window, text+image+video input, and $2.00/$6.00 per-million-token pricing alongside the surrounding catalogue.

פסק הדין

ההפוגה בספטמבר הכריעה את השאלה שכבר הטביעה את ההתמודדות הזו: החצי הזול אינו עוד סיכון נטישה. DeepSeek V4 Pro הוא התשובה של עלות-ביצועים — זול פי שלושה עד שישה בתפוקה, מהיר פי שניים, ברישיון MIT, וכעת מחויב להישאר זמין. Qwen3.8 Max הוא תשובת היכולות — גבוה בארבע נקודות מדד עצמאיות, טענות ארגוניות חזקות יותר, קלט רב-מודאלי, ורישיון מסחרי שמונע ממנו להיכנס לסטאק שלך. בחר בארבע הנקודות כשההערכות שלך מוכיחות שהן חשובות; בחר במחיר כשהחשבון שלך הוא ההערכה שבאמת מכריעה.

שלח את התעבורה עתירת הטוקנים או הרגישה להשהיה אל DeepSeek V4 Pro ואת הסקה המולטימודלית או בעלת הסיכון הגבוה ביותר אל Qwen3.8 Max מתוך אותה אינטגרציה — וזה בדיוק האופן שבו רוב הצוותים משתמשים בסופו של דבר בצירוף הזה: שניהם, כל אחד למה שהוא עושה הכי טוב.

השוואות במאמר הזה1

זוהה מתוך המאמר הזה · בנצ'מרקים: Artificial Analysis · מתעדכן יומית