Zhipu AI logo

GLM-5.3 Flash

Multimodal
Zhipu AI

GLM-5.3 Flash is Zhipu AI's open-weights multimodal mixture-of-experts model released August 2026, the first open-weight release of the glm5_next architecture. A 320B-parameter checkpoint activates only 18B parameters per token and supports a 1M-token context window. It combines sparse and linear attention with an IndexPool mechanism and Manifold-Constrained Hyper-Connections (mHC), was pre-trained on a 30T-token multimodal corpus, and is licensed under MIT. Zhipu says production serving runs on domestically developed AI accelerators with W8A8 quantization.

Key Specifications

Parameters
320.0B
Context
1.0M
Release Date
August 26, 2026
Average Score
64.1%

Timeline

Key dates in the model's history
Announcement
August 26, 2026
Last Update
August 28, 2026
Today
September 8, 2026

Technical Specifications

Parameters
320.0B
Training Tokens
30.0T tokens
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$0.15
Output (per 1M tokens)
$0.50
Max Input Tokens
1.0M
Max Output Tokens
131.1K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

Other Tests

Specialized benchmarks
Terminal-Bench 2.1
Claude Code 2.1.207; max reasoning effortSelf-reported
84.3%
DeepSWE 1.1
mini-swe-agent harnessSelf-reported
63.4%
Toolathlon Verified
Official Toolathlon evaluator, Verified subsetSelf-reported
78.4%
AutomationBench v1.0.6
Official ALE evaluator, v1.0.6Self-reported
48.8%
Agents' Last Exam
Agentic exam, official evaluatorSelf-reported
26.3%
HLE w/ Tools
HLE w/ tools, judged by GPT-5.6-luna (medium)Self-reported
55.3%
Artificial Analysis
Artificial Analysis Intelligence Index v4.1.1. Score 57.Self-reported
57.0%
BabyVision
temp=1.0 top_p=0.95; 164K context; shorter side ≥1.5K pxSelf-reported
53.4%
Chartography
With toolsSelf-reported
78.0%
CharXiv-R
CharXiv Reasoning w/ Tools; temp=1.0 top_p=0.95; 256K contextSelf-reported
89.4%
GDPval-AA
GDPval-AA v2 Elo (max_score 3000); evaluated by Artificial AnalysisSelf-reported
59.1%
Humanity's Last Exam
With tools; temp=1.0 top_p=0.95; 300K context; GPT-5.6-luna (medium) judgeSelf-reported
55.3%
MMVU
Native video input; temp=1.0 top_p=0.95; 256K contextSelf-reported
80.5%
MVBench
Native video input; temp=1.0 top_p=0.95; 256K contextSelf-reported
77.8%
NL2Repo
1M context; rule-based and LLM anti-hacking checksSelf-reported
56.3%
OfficeQA Pro
Treasury Bulletin PDF; no embedded text; temp=1.0; 512K contextSelf-reported
62.4%

License & Metadata

License
mit
Announcement Date
August 26, 2026
Last Updated
August 28, 2026

Compare GLM-5.3 Flash

All comparisons

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.