Gemini 3.1 Flash-Lite
Google's most cost-efficient Gemini 3 model with 1M context, multimodal input, and 2.5x faster time-to-first-token than Gemini 2.5 Flash
Gemini 3.1 Flash-Lite
Google • March 2026
Training Data
Up to early 2025
Gemini 3.1 Flash-Lite
March 2026
Parameters
Not disclosed
Training Method
Multimodal pre-training with RLHF
Context Window
1,000,000 tokens
Knowledge Cutoff
January 2025
Key Features
1M Context Window • Multimodal Input • Extended Thinking
Capabilities
Speed: Outstanding
Cost Efficiency: Outstanding
Multimodal: Good
What's New in This Version
2.5x faster time-to-first-token and 45% faster output generation than Gemini 2.5 Flash at the lowest cost in the Gemini 3 family
Google's most cost-efficient Gemini 3 model with 1M context, multimodal input, and 2.5x faster time-to-first-token than Gemini 2.5 Flash
What's New in This Version
2.5x faster time-to-first-token and 45% faster output generation than Gemini 2.5 Flash at the lowest cost in the Gemini 3 family
Technical Specifications
Key Features
Capabilities
Other Google Models
Explore more models from Google
Gemini 3.7 Flash
Google's most intelligent workhorse model for coding and agents, refining Gemini 3.6 Flash with algorithmic improvements to its core reasoning foundation
Gemini 3.6 Flash
Google's workhorse model delivering better coding, knowledge work, and multimodal performance with ~17% fewer output tokens than Gemini 3.5 Flash
Gemini 3.5 Flash
Google's frontier Flash model built for agentic workflows, coding, and long-horizon tasks at high speed