Google logo

Gemini 3.6 Flash

Multimodal
Google

Gemini 3.6 Flash is Google's fast reasoning model released in July 2026, ranking first on MLE-Bench and delivering strong agentic coding and terminal task performance. It pairs a 1M token input context with up to 65.5K output tokens at $1.50 per million input tokens.

Key Specifications

Parameters
-
Context
1.0M
Release Date
July 21, 2026
Average Score
78.5%

Timeline

Key dates in the model's history
Announcement
July 21, 2026
Last Update
August 27, 2026

Technical Specifications

Parameters
-
Training Tokens
-
Knowledge Cutoff
March 31, 2026
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$1.50
Output (per 1M tokens)
$7.50
Max Input Tokens
1.0M
Max Output Tokens
65.5K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

Other Tests

Specialized benchmarks
CharXiv-R
Search and code executionSelf-reported
89.0%
OSWorld-Verified
Five-run average, single attemptSelf-reported
83.0%
Terminal-Bench 2.1
Terminus-2 harnessSelf-reported
78.0%
MLE-Bench
Partial 30, average position scoreSelf-reported
64.0%

License & Metadata

License
proprietary
Announcement Date
July 21, 2026
Last Updated
August 27, 2026

Compare Gemini 3.6 Flash

All comparisons

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.