Gemini 3.7 Flash vs Kimi K3: Specs & Benchmark Comparison

Gemini 3.7 Flash is developed by Google, while Kimi K3 comes from Moonshot AI. Kimi K3 was released in July 2026, and Gemini 3.7 Flash followed a month later in August 2026. Kimi K3 has a published size of about 2.8 trillion parameters; Google has not disclosed the parameter count of Gemini 3.7 Flash.

We have 4 benchmark results for Gemini 3.7 Flash and 3 benchmark results for Kimi K3, but they were measured on different benchmarks, so there is no like-for-like scoreboard.

Gemini 3.7 Flash is the cheaper API at $0.75 per million input tokens and $3.75 per million output tokens, roughly 4 times cheaper than Kimi K3 at $3 and $15. Both accept a context window of about 1M tokens. Both accept text and images as input.

CharacteristicGemini 3.7 FlashKimi K3
CompanyGoogleMoonshot AI
Release DateAugust 13, 2026July 16, 2026
Parameters2.8T
MultimodalYesYes
Context (input)1.0M1.0M
Context (output)66K1.0M
Input Price / 1M$0.75$3.00
Output Price / 1M$3.75$15.00
Average Score79.6%95.7%

Verdict

Gemini 3.7 Flash leads in 3 out of 4 comparison categories.

Overall Performance

Both models show comparable average scores: Gemini 3.7 Flash — 0.8, Kimi K3 — 1.0.

API Cost

Gemini 3.7 Flash is 4.0x cheaper: input $0.75/1M vs $3.00/1M tokens.

Context Window

Gemini 3.7 Flash supports a larger context: 1M vs 1M tokens.

Recency

Gemini 3.7 Flash is newer: released 8/13/2026 vs 7/16/2026.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — Gemini 3.7 Flash or Kimi K3?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — Gemini 3.7 Flash or Kimi K3?
Gemini 3.7 Flash is cheaper for input: $0.75 per 1M tokens vs $3.00.
Which has a larger context window — Gemini 3.7 Flash or Kimi K3?
Gemini 3.7 Flash supports a larger context: 1,048,576 tokens vs 1,000,000.

The Gemini 3.7 Flash and Kimi K3 comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the Gemini 3.7 Flash or Kimi K3 page. See also the complete list of AI model comparisons.