Kimi K3 vs Qwen3.8 Flash: Specs & Benchmark Comparison

Kimi K3 is developed by Moonshot AI, while Qwen3.8 Flash comes from Alibaba. Kimi K3 was released in July 2026, and Qwen3.8 Flash followed a month later in August 2026. Kimi K3 is the larger model at roughly 2.8 trillion parameters, against 125 billion for Qwen3.8 Flash.

The two models share 2 published benchmarks. Kimi K3 leads on 2 of them. The widest gaps are on MathVision, where Kimi K3 scores 98.0% against 95.7%; GPQA, where Kimi K3 scores 94.0% against 91.7%. Averaged across everything we track, Kimi K3 sits at 95.7% and Qwen3.8 Flash at 68.5%.

Qwen3.8 Flash is the cheaper API at $0.15 per million input tokens and $0.47 per million output tokens, roughly 20 times cheaper than Kimi K3 at $3 and $15. Both accept a context window of about 1M tokens. Kimi K3 accepts text and images as input, while Qwen3.8 Flash accepts text, images, and video. On tooling, only Kimi K3 supports function calling and only Kimi K3 offers structured output.

CharacteristicKimi K3Qwen3.8 Flash
CompanyMoonshot AIAlibaba
Release DateJuly 16, 2026August 26, 2026
Parameters2.8T125B
MultimodalYesYes
Context (input)1.0M1.0M
Context (output)1.0M131K
Input Price / 1M$3.00$0.15
Output Price / 1M$15.00$0.47
Average Score95.7%68.5%
Benchmarks
MathVision98.0%95.7%
GPQA94.0%91.7%

Visual Benchmark Comparison

Kimi K3
Qwen3.8 Flash
MathVision1.0 vs 1.0
1.0
1.0
GPQA0.9 vs 0.9
0.9
0.9

Verdict

Qwen3.8 Flash leads in 3 out of 4 comparison categories.

Overall Performance

Both models show comparable average scores: Kimi K3 — 1.0, Qwen3.8 Flash — 0.7.

API Cost

Qwen3.8 Flash is 29.0x cheaper: input $0.15/1M vs $3.00/1M tokens.

Context Window

Qwen3.8 Flash supports a larger context: 1M vs 1M tokens.

Recency

Qwen3.8 Flash is newer: released 8/26/2026 vs 7/16/2026.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — Kimi K3 or Qwen3.8 Flash?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — Kimi K3 or Qwen3.8 Flash?
Qwen3.8 Flash is cheaper for input: $0.15 per 1M tokens vs $3.00.
Which has a larger context window — Kimi K3 or Qwen3.8 Flash?
Qwen3.8 Flash supports a larger context: 1,048,576 tokens vs 1,000,000.

The Kimi K3 and Qwen3.8 Flash comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the Kimi K3 or Qwen3.8 Flash page. See also the complete list of AI model comparisons.