UseRightAI
HomeModelsAsk AIComparePricingWhat's New
UseRightAICut through AI hype. Pick what works.

Independent AI model tracker. Live pricing, real benchmarks, zero vendor bias.

X (Twitter)LinkedInUpdatesContact

Compare

Opus 4.8 vs Opus 4.7Fable 5 vs Opus 4.8New AI Models 2026ChatGPT vs ClaudeGPT-4o vs Claude SonnetClaude vs GeminiDeepSeek vs ChatGPTMistral vs ClaudeGemini Flash vs GPT-4o MiniLlama vs ChatGPTAll comparisons →Build your own →

Best For

CodingWritingDevelopersProduct ManagersDesignersSalesBest Cheap AIBest Free AI

Pricing & Data

API Token PricingCost per TaskPrice HistoryBenchmark ScoresPrivacy & SafetySubscription PlansCost CalculatorWhich AI is Cheapest?Cheapest AI APIs

Company

About UseRightAIContactWhat ChangedAll ModelsEditorial PolicyDisclosuresPrivacy PolicyTerms of Service

© 2026 UseRightAI. Independent · Free forever · Not affiliated with any AI provider.

Affiliate links are clearly labeled. See disclosures.

Hangzhou, China · Founded 2023 (Qwen team, Alibaba Cloud)

Alibaba Qwen

The Qwen family — frontier scale at Chinese-cloud prices.

Alibaba's Qwen team ships some of the strongest models outside the US labs. Qwen 3.8 Max beat GPT-5.6 Sol on SWE-bench Pro (67.7 vs 64.6) at launch, and Qwen 3.8 Flash delivers SWE-bench Pro 62.5 from just 6B active parameters at $0.16/1M.

Rankings refresh dailyScored on 6 criteriaNo paid rankings
  • Qwen 3.8 Max beat GPT-5.6 Sol on SWE-bench Pro at launch (67.7 vs 64.6)
  • Qwen 3.8 Flash: SWE-bench Pro 62.5 at $0.16/1M — from 6B active parameters
  • Qwen3.8-Flash-Next open weights preview the Qwen4 architecture
3 models

All Alibaba Qwen Models

Every Alibaba Qwen model in the directory, ranked by overall capability score.

AlibabaBalanced

Qwen 3.8 Max

Alibaba's largest model ever — a 2.4-trillion-parameter MoE (95B active) multimodal flagship that beat GPT-5.6 Sol on SWE-bench Pro and ranks #2 globally for vision.

Verdict
Best Chinese flagship — beats GPT-5.6 Sol on coding, #2 globally for vision.
Quality score
90%
Pricing
$2.00/1M in
$6.00/1M out
Speed
Balanced
3/5 speed
Context
1M tokens
Announced August 3, 2026 on Alibaba Cloud Model Studio; open weights promised a week after launch. $2/$6 is first-party Model Studio pricing; cache reads from $0.17/1M. Announcement moved Alibaba stock +7% in Hong Kong.
Open weightsMultimodalVisionCoding1M context
Best for
Multimodal and vision-heavy workloads at scale
View model
AlibabaBalanced

Qwen 3.7 Max

Alibaba's first closed-weight flagship — an agent-first model with native extended thinking, built to run autonomously for up to ~35 hours firing thousands of tool calls.

Verdict
Agent-first Qwen flagship, superseded by Qwen 3.8 Max.
Quality score
87%
Pricing
$2.50/1M in
$7.50/1M out
Speed
Balanced
3/5 speed
Context
1M tokens
Announced at Alibaba Cloud Summit May 20, 2026. API-only on DashScope/Model Studio. Cached input $0.25/1M.
AgenticReasoningLong contextClosed weights
Best for
Long-horizon autonomous agent runs
View model
AlibabaBudget

Qwen 3.8 Flash

Alibaba's preview of the Qwen4 architecture — 125B parameters with only 6B active per token, at sixteen cents per million input.

Verdict
SWE-bench Pro 62.5 at sixteen cents per million input.
Quality score
77%
Pricing
$0.16/1M in
$0.47/1M out
Speed
Very fast
5/5 speed
Context
991k tokens
Released August 26, 2026. The open-weight release is Qwen3.8-Flash-Next, a preview of the Qwen4 architecture: 125B mixture-of-experts with 6B active per token, a 51B n-gram embedding table and a 4B multi-token prediction layer. Qwen 3.8 Flash is the production API version on Qwen Cloud at $0.16/$0.47.
Open weightsBudgetCodingLong context
Best for
Cheap high-throughput coding and reasoning
View model

Alibaba Qwen API Pricing

Per 1 million tokens. Updated when providers change prices.

ModelInput / 1MOutput / 1MContextSpeed
Qwen 3.8 Max
Balanced
$2.00/1M$6.00/1M1MBalanced
Qwen 3.7 Max
Balanced
$2.50/1M$7.50/1M1MBalanced
Qwen 3.8 Flash
Budget
$0.16/1M$0.47/1M991KVery fast
Compare all providers →

Compare Alibaba Qwen Models

Head-to-head comparisons for the most-searched questions.

Qwen 3.8 Max vs Qwen 3.8 FlashOpen compare tool →

Go deeper on Alibaba Qwen

Alibaba · Coding
Best Alibaba Model for CodingEvery Alibaba model ranked for coding — capability scores, price per 1M tokens, and context windows, with a clear top pick and a budget option.Read guide
Alibaba · Writing
Best Alibaba Model for WritingEvery Alibaba model ranked for writing — capability scores, price per 1M tokens, and context windows, with a clear top pick and a budget option.Read guide
Alibaba · Research
Best Alibaba Model for ResearchEvery Alibaba model ranked for research — capability scores, price per 1M tokens, and context windows, with a clear top pick and a budget option.Read guide
Alibaba · Long Context
Best Alibaba Model for Long ContextEvery Alibaba model ranked for long-context work — capability scores, price per 1M tokens, and context windows, with a clear top pick and a budget option.Read guide
Alibaba · Pricing
Cheapest Alibaba Model Worth UsingEvery Alibaba model ranked by API price — what the cheapest option costs per 1M tokens, what you give up, and which budget pick is actually worth using.Read guide
Alternatives
Best Qwen AlternativesThe best Qwen alternatives in 2026, compared on real capability scores, API pricing, and context windows — with free and open-weight options included.Read guide

Newsletter

Get notified when Alibaba Qwen releases new models

Pricing changes, new releases, and ranking shifts — straight to your inbox.

No spam. Useful updates only. Affiliate disclosures always clearly labeled.

Alibaba Qwen FAQ

What is Alibaba's best AI model in 2026?

Qwen 3.8 Max ($2/$6) is Alibaba's flagship — a 2.4T-parameter MoE that beat GPT-5.6 Sol on SWE-bench Pro at launch and ranks #2 globally on vision. Qwen 3.8 Flash ($0.16/$0.47) is the value pick: SWE-bench Pro 62.5 from a 125B MoE with only 6B active parameters per token.

Are Qwen models open source?

Partially. Qwen 3.8 Max is API-only — a break from Qwen tradition. But the Qwen3.8-Flash-Next release (August 26, 2026) is open-weight and previews the Qwen4 architecture: 125B mixture-of-experts, 6B active per token, plus a 51B n-gram embedding table.

How does Qwen compare to Claude and GPT?

Qwen 3.8 Max genuinely beat GPT-5.6 Sol on SWE-bench Pro at launch (67.7 vs 64.6), though the Claude frontier — Fable 5 at 80.3% — remains well ahead. Where Qwen wins is price: Max costs a third of Sol, and Qwen 3.8 Flash delivers GLM-5.2-class coding at $0.16/1M input.

Where can I access Qwen models?

Via Alibaba Cloud's Model Studio API (international endpoint available), chat.qwen.ai for consumer use, and — for the open-weight Flash-Next variant — self-hosting or third-party hosts. Data-residency-sensitive teams should note the first-party API routes through Alibaba Cloud.

Explore other providers

OpenAIAnthropicGooglexAIMetaMistralDeepSeekBrowse all models →