Gemini 3.7 Flash vs Kimi K3: Specs & Benchmark Comparison
Gemini 3.7 Flash is developed by Google, while Kimi K3 comes from Moonshot AI. Kimi K3 was released in July 2026, and Gemini 3.7 Flash followed a month later in August 2026. Kimi K3 has a published size of about 2.8 trillion parameters; Google has not disclosed the parameter count of Gemini 3.7 Flash.
We have 4 benchmark results for Gemini 3.7 Flash and 3 benchmark results for Kimi K3, but they were measured on different benchmarks, so there is no like-for-like scoreboard.
Gemini 3.7 Flash is the cheaper API at $0.75 per million input tokens and $3.75 per million output tokens, roughly 4 times cheaper than Kimi K3 at $3 and $15. Both accept a context window of about 1M tokens. Both accept text and images as input.
| Characteristic | Gemini 3.7 Flash | Kimi K3 |
|---|---|---|
| Company | Moonshot AI | |
| Release Date | August 13, 2026 | July 16, 2026 |
| Parameters | — | 2.8T |
| Multimodal | Yes | Yes |
| Context (input) | 1.0M | 1.0M |
| Context (output) | 66K | 1.0M |
| Input Price / 1M | $0.75 | $3.00 |
| Output Price / 1M | $3.75 | $15.00 |
| Average Score | 79.6% | 95.7% |
Verdict
Gemini 3.7 Flash leads in 3 out of 4 comparison categories.
Both models show comparable average scores: Gemini 3.7 Flash — 0.8, Kimi K3 — 1.0.
Gemini 3.7 Flash is 4.0x cheaper: input $0.75/1M vs $3.00/1M tokens.
Gemini 3.7 Flash supports a larger context: 1M vs 1M tokens.
Gemini 3.7 Flash is newer: released 8/13/2026 vs 7/16/2026.
More About These Models
Related Comparisons
Frequently Asked Questions
Which is better for coding — Gemini 3.7 Flash or Kimi K3?
Which model is cheaper — Gemini 3.7 Flash or Kimi K3?
Which has a larger context window — Gemini 3.7 Flash or Kimi K3?
The Gemini 3.7 Flash and Kimi K3 comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the Gemini 3.7 Flash or Kimi K3 page. See also the complete list of AI model comparisons.