Gemini 3.7 Flash vs Qwen3.8 Flash: Specs & Benchmark Comparison

Gemini 3.7 Flash is developed by Google, while Qwen3.8 Flash comes from Alibaba. Both were released in August 2026. Qwen3.8 Flash has a published size of about 125 billion parameters; Google has not disclosed the parameter count of Gemini 3.7 Flash.

The two models share 5 published benchmarks. Gemini 3.7 Flash leads on 3 of them, Qwen3.8 Flash on 2. The widest gaps are on OSWorld 2.0, where Gemini 3.7 Flash scores 47.9% against 19.4%; Agents' Last Exam, where Qwen3.8 Flash scores 51.2% against 26.3%. Averaged across everything we track, Gemini 3.7 Flash sits at 59.8% and Qwen3.8 Flash at 68.5%.

Qwen3.8 Flash is the cheaper API at $0.15 per million input tokens and $0.47 per million output tokens, roughly 5 times cheaper than Gemini 3.7 Flash at $0.75 and $3.75. Both accept a context window of about 1M tokens. Gemini 3.7 Flash accepts text and images as input, while Qwen3.8 Flash accepts text, images, and video. On tooling, only Gemini 3.7 Flash supports function calling and only Gemini 3.7 Flash offers structured output.

CharacteristicGemini 3.7 FlashQwen3.8 Flash
CompanyGoogleAlibaba
Release DateAugust 13, 2026August 26, 2026
Parameters125B
MultimodalYesYes
Context (input)1.0M1.0M
Context (output)66K131K
Input Price / 1M$0.75$0.15
Output Price / 1M$3.75$0.47
Average Score59.8%68.5%
Benchmarks
OSWorld 2.047.9%19.4%
Agents' Last Exam26.3%51.2%
LVBench85.4%76.6%
DeepSWE 1.165.3%58.7%
CharXiv-R88.7%90.6%

Visual Benchmark Comparison

Gemini 3.7 Flash
Qwen3.8 Flash
OSWorld 2.00.5 vs 0.2
0.5
0.2
Agents' Last Exam0.3 vs 0.5
0.3
0.5
LVBench0.9 vs 0.8
0.9
0.8
DeepSWE 1.10.7 vs 0.6
0.7
0.6
CharXiv-R0.9 vs 0.9
0.9
0.9

Verdict

Qwen3.8 Flash leads in 1 out of 4 comparison categories.

Overall Performance

Both models show comparable average scores: Gemini 3.7 Flash — 0.6, Qwen3.8 Flash — 0.7.

API Cost

Qwen3.8 Flash is 7.3x cheaper: input $0.15/1M vs $0.75/1M tokens.

Context Window

Same context size: 1M tokens.

Recency

Both models were released around the same time: 8/13/2026 and 8/26/2026.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — Gemini 3.7 Flash or Qwen3.8 Flash?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — Gemini 3.7 Flash or Qwen3.8 Flash?
Qwen3.8 Flash is cheaper for input: $0.15 per 1M tokens vs $0.75.
Which has a larger context window — Gemini 3.7 Flash or Qwen3.8 Flash?
Gemini 3.7 Flash supports a larger context: 1,048,576 tokens vs 1,048,576.

The Gemini 3.7 Flash and Qwen3.8 Flash comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the Gemini 3.7 Flash or Qwen3.8 Flash page. See also the complete list of AI model comparisons.