GLM-5.3 Flash vs Kimi K3: Specs & Benchmark Comparison

GLM-5.3 Flash is developed by Zhipu AI, while Kimi K3 comes from Moonshot AI. Kimi K3 was released in July 2026, and GLM-5.3 Flash followed a month later in August 2026. Kimi K3 is the larger model at roughly 2.8 trillion parameters, against 320 billion for GLM-5.3 Flash.

The two models share 8 published benchmarks. Kimi K3 leads on 6 of them, GLM-5.3 Flash on 2. The widest gaps are on BabyVision, where Kimi K3 scores 85.7% against 53.4%; AutomationBench v1.0.6, where GLM-5.3 Flash scores 48.8% against 30.8%. Averaged across everything we track, GLM-5.3 Flash sits at 64.1% and Kimi K3 at 67.8%.

GLM-5.3 Flash is the cheaper API at $0.15 per million input tokens and $0.50 per million output tokens, roughly 20 times cheaper than Kimi K3 at $3 and $15. Both accept a context window of about 1M tokens. GLM-5.3 Flash accepts text, images, and video as input, while Kimi K3 accepts text and images.

CharacteristicGLM-5.3 FlashKimi K3
CompanyZhipu AIMoonshot AI
Release DateAugust 26, 2026July 16, 2026
Parameters320B2.8T
MultimodalYesYes
Context (input)1.0M1.0M
Context (output)131K1.0M
Input Price / 1M$0.15$3.00
Output Price / 1M$0.50$15.00
Average Score64.1%67.8%
Benchmarks
BabyVision53.4%85.7%
AutomationBench v1.0.648.8%30.8%
DeepSWE 1.163.4%69.0%
Toolathlon Verified78.4%73.2%
Terminal-Bench 2.184.3%88.3%
CharXiv-R89.4%91.3%
OfficeQA Pro62.4%63.3%
Humanity's Last Exam55.3%56.0%

Visual Benchmark Comparison

GLM-5.3 Flash
Kimi K3
BabyVision0.5 vs 0.9
0.5
0.9
AutomationBench v1.0.60.5 vs 0.3
0.5
0.3
DeepSWE 1.10.6 vs 0.7
0.6
0.7
Toolathlon Verified0.8 vs 0.7
0.8
0.7
Terminal-Bench 2.10.8 vs 0.9
0.8
0.9
CharXiv-R0.9 vs 0.9
0.9
0.9
OfficeQA Pro0.6 vs 0.6
0.6
0.6
Humanity's Last Exam0.6 vs 0.6
0.6
0.6

Verdict

GLM-5.3 Flash leads in 3 out of 4 comparison categories.

Overall Performance

Both models show comparable average scores: GLM-5.3 Flash — 0.6, Kimi K3 — 0.7.

API Cost

GLM-5.3 Flash is 27.7x cheaper: input $0.15/1M vs $3.00/1M tokens.

Context Window

GLM-5.3 Flash supports a larger context: 1M vs 1M tokens.

Recency

GLM-5.3 Flash is newer: released 8/26/2026 vs 7/16/2026.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — GLM-5.3 Flash or Kimi K3?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — GLM-5.3 Flash or Kimi K3?
GLM-5.3 Flash is cheaper for input: $0.15 per 1M tokens vs $3.00.
Which has a larger context window — GLM-5.3 Flash or Kimi K3?
GLM-5.3 Flash supports a larger context: 1,048,576 tokens vs 1,000,000.

The GLM-5.3 Flash and Kimi K3 comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the GLM-5.3 Flash or Kimi K3 page. See also the complete list of AI model comparisons.