Qwen3.5 27B vs Qwen3.8 Max: Specs & Benchmark Comparison

Qwen3.5 27B and Qwen3.8 Max both come from Alibaba, Qwen3.5 27B was released in March 2026, and Qwen3.8 Max followed 5 months later in August 2026. Qwen3.8 Max is the larger model at roughly 2.4 trillion parameters, against 27 billion for Qwen3.5 27B.

The two models share 12 published benchmarks. Qwen3.8 Max leads on 11 of them, Qwen3.5 27B on 1. The widest gaps are on OSWorld-Verified, where Qwen3.8 Max scores 86.1% against 56.2%; WideSearch, where Qwen3.8 Max scores 81.9% against 61.1%. Averaged across everything we track, Qwen3.5 27B sits at 70.3% and Qwen3.8 Max at 71.0%.

Qwen3.8 Max is available through an API at $2.50 per million input tokens with a 1M-token context window. We do not track a hosted API for Qwen3.5 27B, so pricing and context cannot be compared directly.

CharacteristicQwen3.5 27BQwen3.8 Max
CompanyAlibabaAlibaba
Release DateMarch 1, 2026August 2, 2026
Parameters27B2.4T
MultimodalNoYes
Context (input)1.0M
Context (output)131K
Input Price / 1M$2.50
Output Price / 1M$6.25
Average Score70.3%71.0%
Benchmarks
OSWorld-Verified56.2%86.1%
WideSearch61.1%81.9%
ERQA60.5%77.8%
ScreenSpot Pro70.3%84.5%
LVBench73.6%81.8%
MMMU-Pro75.0%82.3%
GPQA85.5%92.6%
IFBench76.5%82.8%
LongBench v260.6%66.3%
Humanity's Last Exam48.5%43.6%
RealWorldQA83.7%88.0%
VideoMME w sub.87.0%90.4%

Visual Benchmark Comparison

Qwen3.5 27B
Qwen3.8 Max
OSWorld-Verified0.6 vs 0.9
0.6
0.9
WideSearch0.6 vs 0.8
0.6
0.8
ERQA0.6 vs 0.8
0.6
0.8
ScreenSpot Pro0.7 vs 0.8
0.7
0.8
LVBench0.7 vs 0.8
0.7
0.8
MMMU-Pro0.8 vs 0.8
0.8
0.8
GPQA0.9 vs 0.9
0.9
0.9
IFBench0.8 vs 0.8
0.8
0.8
LongBench v20.6 vs 0.7
0.6
0.7
Humanity's Last Exam0.5 vs 0.4
0.5
0.4
RealWorldQA0.8 vs 0.9
0.8
0.9
VideoMME w sub.0.9 vs 0.9
0.9
0.9

Verdict

Qwen3.8 Max leads in 1 out of 2 comparison categories.

Overall Performance

Both models show comparable average scores: Qwen3.5 27B — 0.7, Qwen3.8 Max — 0.7.

Recency

Qwen3.8 Max is newer: released 8/2/2026 vs 3/1/2026.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — Qwen3.5 27B or Qwen3.8 Max?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — Qwen3.5 27B or Qwen3.8 Max?
API pricing data is available on the individual model pages.
Which has a larger context window — Qwen3.5 27B or Qwen3.8 Max?
Context window data is available on the individual model pages.

The Qwen3.5 27B and Qwen3.8 Max comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the Qwen3.5 27B or Qwen3.8 Max page. See also the complete list of AI model comparisons.