Qwen 3.8 Max
Alibaba's largest model ever — a 2.4-trillion-parameter MoE (95B active) multimodal flagship that beat GPT-5.6 Sol on SWE-bench Pro and ranks #2 globally for vision.
The Qwen family — frontier scale at Chinese-cloud prices.
Alibaba's Qwen team ships some of the strongest models outside the US labs. Qwen 3.8 Max beat GPT-5.6 Sol on SWE-bench Pro (67.7 vs 64.6) at launch, and Qwen 3.8 Flash delivers SWE-bench Pro 62.5 from just 6B active parameters at $0.16/1M.
Every Alibaba Qwen model in the directory, ranked by overall capability score.
Alibaba's largest model ever — a 2.4-trillion-parameter MoE (95B active) multimodal flagship that beat GPT-5.6 Sol on SWE-bench Pro and ranks #2 globally for vision.
Alibaba's first closed-weight flagship — an agent-first model with native extended thinking, built to run autonomously for up to ~35 hours firing thousands of tool calls.
Alibaba's preview of the Qwen4 architecture — 125B parameters with only 6B active per token, at sixteen cents per million input.
Per 1 million tokens. Updated when providers change prices.
| Model | Input / 1M | Output / 1M | Context | Speed |
|---|---|---|---|---|
| Qwen 3.8 Max Balanced | $2.00/1M | $6.00/1M | 1M | Balanced |
| Qwen 3.7 Max Balanced | $2.50/1M | $7.50/1M | 1M | Balanced |
| Qwen 3.8 Flash Budget | $0.16/1M | $0.47/1M | 991K | Very fast |
Head-to-head comparisons for the most-searched questions.
Newsletter
Pricing changes, new releases, and ranking shifts — straight to your inbox.
No spam. Useful updates only. Affiliate disclosures always clearly labeled.
Qwen 3.8 Max ($2/$6) is Alibaba's flagship — a 2.4T-parameter MoE that beat GPT-5.6 Sol on SWE-bench Pro at launch and ranks #2 globally on vision. Qwen 3.8 Flash ($0.16/$0.47) is the value pick: SWE-bench Pro 62.5 from a 125B MoE with only 6B active parameters per token.
Partially. Qwen 3.8 Max is API-only — a break from Qwen tradition. But the Qwen3.8-Flash-Next release (August 26, 2026) is open-weight and previews the Qwen4 architecture: 125B mixture-of-experts, 6B active per token, plus a 51B n-gram embedding table.
Qwen 3.8 Max genuinely beat GPT-5.6 Sol on SWE-bench Pro at launch (67.7 vs 64.6), though the Claude frontier — Fable 5 at 80.3% — remains well ahead. Where Qwen wins is price: Max costs a third of Sol, and Qwen 3.8 Flash delivers GLM-5.2-class coding at $0.16/1M input.
Via Alibaba Cloud's Model Studio API (international endpoint available), chat.qwen.ai for consumer use, and — for the open-weight Flash-Next variant — self-hosting or third-party hosts. Data-residency-sensitive teams should note the first-party API routes through Alibaba Cloud.