Google logo

Gemini 3.5 Flash-Lite

Multimodal
Google

Gemini 3.5 Flash-Lite is Google's most cost-efficient Gemini 3.5 model, designed for high-volume, latency-sensitive applications at a starting price of $0.30 per million input tokens. It offers a 1M token input context window with up to 65.5K output tokens while retaining multimodal input support.

Key Specifications

Parameters
-
Context
1.0M
Release Date
July 21, 2026
Average Score
-

Timeline

Key dates in the model's history
Announcement
July 21, 2026
Last Update
August 27, 2026

Technical Specifications

Parameters
-
Training Tokens
-
Knowledge Cutoff
March 31, 2026
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$0.30
Output (per 1M tokens)
$2.50
Max Input Tokens
1.0M
Max Output Tokens
65.5K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

License & Metadata

License
proprietary
Announcement Date
July 21, 2026
Last Updated
August 27, 2026

Compare Gemini 3.5 Flash-Lite

All comparisons

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.