Google logo

Gemini 3.7 Flash

Multimodal
Google

Gemini 3.7 Flash is Google's flagship fast model released in August 2026, ranking first on long-context, MRCR v2 (8-needle), WebDev Arena, and Harvey LAB-AA benchmarks. It combines a 1M token context window with top-tier retrieval, reasoning, and web development performance at $0.75 per million input tokens.

Key Specifications

Parameters
-
Context
1.0M
Release Date
August 13, 2026
Average Score
79.6%

Timeline

Key dates in the model's history
Announcement
August 13, 2026
Last Update
August 27, 2026

Technical Specifications

Parameters
-
Training Tokens
-
Knowledge Cutoff
March 31, 2026
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$0.75
Output (per 1M tokens)
$3.75
Max Input Tokens
1.0M
Max Output Tokens
65.5K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

Other Tests

Specialized benchmarks
WebDev Arena
Arena.ai WebDev Arena Elo (blog); model card lists the same Elo under Code Arena.Self-reported
79.4%
GDPval-AA
GDPval-AA v2 Elo (max_score 3000).Self-reported
50.8%
MRCR v2 (8-needle)
128k cumulative averageSelf-reported
97.0%
Harvey LAB-AA
Harvey LAB-AA; complex legal workflows.Self-reported
91.0%

License & Metadata

License
proprietary
Announcement Date
August 13, 2026
Last Updated
August 27, 2026

Compare Gemini 3.7 Flash

All comparisons

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.