GLM-5.2 vs Qwen3.6 Plus: Specs & Benchmark Comparison

GLM-5.2 is developed by Zhipu AI, while Qwen3.6 Plus comes from Alibaba. Qwen3.6 Plus was released in March 2026, and GLM-5.2 followed 3 months later in June 2026. GLM-5.2 has a published size of about 753 billion parameters; Alibaba has not disclosed the parameter count of Qwen3.6 Plus.

The two models share 11 published benchmarks. GLM-5.2 leads on 10 of them, Qwen3.6 Plus on 1. The widest gaps are on FrontierSWE, where GLM-5.2 scores 74.0% against 22.0%; Humanity's Last Exam, where GLM-5.2 scores 54.7% against 28.8%. Averaged across everything we track, GLM-5.2 sits at 60.9% and Qwen3.6 Plus at 74.5%.

Qwen3.6 Plus is the cheaper API at $0.50 per million input tokens and $3 per million output tokens, roughly 3 times cheaper than GLM-5.2 at $1.40 and $4.40. Both accept a context window of about 1M tokens. GLM-5.2 accepts text as input, while Qwen3.6 Plus accepts text and images.

CharacteristicGLM-5.2Qwen3.6 Plus
CompanyZhipu AIAlibaba
Release DateJune 16, 2026March 31, 2026
Parameters753B
MultimodalNoYes
Context (input)1.0M1.0M
Context (output)131K66K
Input Price / 1M$1.40$0.50
Output Price / 1M$4.40$3.00
Average Score60.9%74.5%
Benchmarks
FrontierSWE74.0%22.0%
Humanity's Last Exam54.7%28.8%
NL2Repo48.9%37.9%
Toolathlon48.2%39.8%
IMO-AnswerBench91.0%83.8%
SWE-Bench Pro62.1%56.6%
HMMT Feb 2692.5%87.8%
AIME 202699.0%95.0%
HMMT 202594.0%97.0%
MCP Atlas76.8%74.1%
GPQA91.0%90.4%

Visual Benchmark Comparison

GLM-5.2
Qwen3.6 Plus
FrontierSWE0.7 vs 0.2
0.7
0.2
Humanity's Last Exam0.5 vs 0.3
0.5
0.3
NL2Repo0.5 vs 0.4
0.5
0.4
Toolathlon0.5 vs 0.4
0.5
0.4
IMO-AnswerBench0.9 vs 0.8
0.9
0.8
SWE-Bench Pro0.6 vs 0.6
0.6
0.6
HMMT Feb 260.9 vs 0.9
0.9
0.9
AIME 20261.0 vs 0.9
1.0
0.9
HMMT 20250.9 vs 1.0
0.9
1.0
MCP Atlas0.8 vs 0.7
0.8
0.7
GPQA0.9 vs 0.9
0.9
0.9

Verdict

GLM-5.2 leads in 2 out of 4 comparison categories.

Overall Performance

Both models show comparable average scores: GLM-5.2 — 0.6, Qwen3.6 Plus — 0.7.

API Cost

Qwen3.6 Plus is 1.7x cheaper: input $0.50/1M vs $1.40/1M tokens.

Context Window

GLM-5.2 supports a larger context: 1M vs 1M tokens.

Recency

GLM-5.2 is newer: released 6/16/2026 vs 3/31/2026.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — GLM-5.2 or Qwen3.6 Plus?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — GLM-5.2 or Qwen3.6 Plus?
Qwen3.6 Plus is cheaper for input: $0.50 per 1M tokens vs $1.40.
Which has a larger context window — GLM-5.2 or Qwen3.6 Plus?
GLM-5.2 supports a larger context: 1,048,576 tokens vs 1,000,000.

The GLM-5.2 and Qwen3.6 Plus comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the GLM-5.2 or Qwen3.6 Plus page. See also the complete list of AI model comparisons.