Grok 4.5 vs Kimi K2.6: Specs & Benchmark Comparison

Grok 4.5 is developed by xAI, while Kimi K2.6 comes from Moonshot AI. Kimi K2.6 was released in April 2026, and Grok 4.5 followed 3 months later in July 2026. Kimi K2.6 has a published size of about 1.0 trillion parameters; xAI has not disclosed the parameter count of Grok 4.5.

The two models share 2 published benchmarks. Grok 4.5 leads on 2 of them. The widest gaps are on SWE-Bench Pro, where Grok 4.5 scores 65.0% against 58.6%; GPQA, where Grok 4.5 scores 93.0% against 90.5%. Averaged across everything we track, Grok 4.5 sits at 52.9% and Kimi K2.6 at 71.2%.

Kimi K2.6 is the cheaper API at $1.20 per million input tokens and $4.50 per million output tokens, about 67% below Grok 4.5 at $2 and $6. Grok 4.5 takes the larger context window at 500K tokens, compared with 262K for Kimi K2.6. Both accept text and images as input.

CharacteristicGrok 4.5Kimi K2.6
CompanyxAIMoonshot AI
Release DateJuly 16, 2026April 20, 2026
Parameters1.0T
MultimodalYesYes
Context (input)500K262K
Context (output)131K
Input Price / 1M$2.00$1.20
Output Price / 1M$6.00$4.50
Average Score52.9%71.2%
Benchmarks
SWE-Bench Pro65.0%58.6%
GPQA93.0%90.5%

Visual Benchmark Comparison

Grok 4.5
Kimi K2.6
SWE-Bench Pro0.7 vs 0.6
0.7
0.6
GPQA0.9 vs 0.9
0.9
0.9

Verdict

Grok 4.5 leads in 2 out of 4 comparison categories.

Overall Performance

Both models show comparable average scores: Grok 4.5 — 0.5, Kimi K2.6 — 0.7.

API Cost

Kimi K2.6 is 1.4x cheaper: input $1.20/1M vs $2.00/1M tokens.

Context Window

Grok 4.5 supports a larger context: 500K vs 262K tokens.

Recency

Grok 4.5 is newer: released 7/16/2026 vs 4/20/2026.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — Grok 4.5 or Kimi K2.6?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — Grok 4.5 or Kimi K2.6?
Kimi K2.6 is cheaper for input: $1.20 per 1M tokens vs $2.00.
Which has a larger context window — Grok 4.5 or Kimi K2.6?
Grok 4.5 supports a larger context: 500,000 tokens vs 262,144.

The Grok 4.5 and Kimi K2.6 comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the Grok 4.5 or Kimi K2.6 page. See also the complete list of AI model comparisons.