Gemini 3.5 Flash-Lite
MultimodalGemini 3.5 Flash-Lite is Google's most cost-efficient Gemini 3.5 model, designed for high-volume, latency-sensitive applications at a starting price of $0.30 per million input tokens. It offers a 1M token input context window with up to 65.5K output tokens while retaining multimodal input support.
Key Specifications
Parameters
-
Context
1.0M
Release Date
July 21, 2026
Average Score
-
Timeline
Key dates in the model's history
Announcement
July 21, 2026
Last Update
August 27, 2026
Technical Specifications
Parameters
-
Training Tokens
-
Knowledge Cutoff
March 31, 2026
Family
-
Capabilities
MultimodalZeroEval
Pricing & Availability
Input (per 1M tokens)
$0.30
Output (per 1M tokens)
$2.50
Max Input Tokens
1.0M
Max Output Tokens
65.5K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning
Benchmark Results
Model performance metrics across various tests and benchmarks
License & Metadata
License
proprietary
Announcement Date
July 21, 2026
Last Updated
August 27, 2026
Compare Gemini 3.5 Flash-Lite
All comparisonsSimilar Models
All ModelsGemini 2.5 Flash-Lite
MM
Best score:0.6 (GPQA)
Released:Jun 2025
Price:$0.10/1M tokens
Gemini 2.0 Flash Thinking
MM
Best score:0.7 (GPQA)
Released:Jan 2025
Gemini 1.5 Pro
MM
Best score:0.9 (MMLU)
Released:May 2024
Price:$2.50/1M tokens
Gemini 2.5 Pro
MM
Best score:0.8 (GPQA)
Released:May 2025
Price:$1.25/1M tokens
Gemini 2.5 Pro Preview 06-05
MM
Best score:0.9 (GPQA)
Released:Jun 2025
Price:$1.25/1M tokens
Gemini 2.5 Flash
MM
Best score:0.8 (GPQA)
Released:May 2025
Price:$0.30/1M tokens
Gemini 2.0 Flash-Lite
MM
Best score:0.5 (GPQA)
Released:Feb 2025
Price:$0.07/1M tokens
Gemini 3 Pro
MM
Best score:0.9 (GPQA)
Released:Nov 2025
Price:$2.00/1M tokens
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.