GLM-5.2 vs Inkling-Small: Specs & Benchmark Comparison

GLM-5.2 is developed by Zhipu AI, while Inkling-Small comes from Thinking Machines Lab. GLM-5.2 was released in June 2026, and Inkling-Small followed a month later in July 2026. GLM-5.2 is the larger model at roughly 753 billion parameters, against 276 billion for Inkling-Small.

The two models share 2 published benchmarks. GLM-5.2 leads on 2 of them. The widest gaps are on AIME 2026, where GLM-5.2 scores 99.0% against 95.5%; GPQA, where GLM-5.2 scores 91.0% against 89.5%. Averaged across everything we track, GLM-5.2 sits at 94.3% and Inkling-Small at 60.3%.

Inkling-Small is the cheaper API at $0.30 per million input tokens and $1.20 per million output tokens, roughly 5 times cheaper than GLM-5.2 at $1.40 and $4.40. GLM-5.2 takes the larger context window at 1M tokens, compared with 256K for Inkling-Small. GLM-5.2 accepts text as input, while Inkling-Small accepts text, images, and audio. On tooling, only GLM-5.2 supports function calling and only GLM-5.2 offers structured output.

CharacteristicGLM-5.2Inkling-Small
CompanyZhipu AIThinking Machines Lab
Release DateJune 16, 2026July 30, 2026
Parameters753B276B
MultimodalNoYes
Context (input)1.0M256K
Context (output)131K256K
Input Price / 1M$1.40$0.30
Output Price / 1M$4.40$1.20
Average Score94.3%60.3%
Benchmarks
AIME 202699.0%95.5%
GPQA91.0%89.5%

Visual Benchmark Comparison

GLM-5.2
Inkling-Small
AIME 20261.0 vs 1.0
1.0
1.0
GPQA0.9 vs 0.9
0.9
0.9

Verdict

Inkling-Small leads in 2 out of 4 comparison categories.

Overall Performance

Both models show comparable average scores: GLM-5.2 — 0.9, Inkling-Small — 0.6.

API Cost

Inkling-Small is 3.9x cheaper: input $0.30/1M vs $1.40/1M tokens.

Context Window

GLM-5.2 supports a larger context: 1M vs 256K tokens.

Recency

Inkling-Small is newer: released 7/30/2026 vs 6/16/2026.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — GLM-5.2 or Inkling-Small?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — GLM-5.2 or Inkling-Small?
Inkling-Small is cheaper for input: $0.30 per 1M tokens vs $1.40.
Which has a larger context window — GLM-5.2 or Inkling-Small?
GLM-5.2 supports a larger context: 1,048,576 tokens vs 256,000.

The GLM-5.2 and Inkling-Small comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the GLM-5.2 or Inkling-Small page. See also the complete list of AI model comparisons.