Gemini 3.5 Flash-Lite
Gemini 3.5 Flash-Lite is a lightweight language model released by Google as part of its Gemini family, optimized for efficient AI agent deployment. It achieves 350 output tokens per second with pricing of $0.30 per million input tokens and $2.50 per million output tokens, enabling developers to build cost-effective agentic workflows. The model was released alongside Gemini 3.6 Flash and 3.5 Flash Cyber to provide lower-latency, more affordable alternatives for production workloads.
Updated 3 August 2026
Specifications
No specifications recorded yet.
Latest developments
Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
TLDR Dev · 1 week ago ·
38
Google Releases Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber: A Cheaper, More Token-Efficient Flash Tier Built for Agentic Workloads
MarkTechPost · 1 week ago ·
41
Google releases three new Gemini models — but no 3.5 Pro
TechCrunch AI · 1 week ago ·
44
Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
Google DeepMind · 1 week ago ·
9
Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
Google DeepMind · 1 week ago ·
12
Google ships 3 new Gemini models. Just not the one everyone’s waiting for.
The New Stack · 1 week ago ·
21
July 2026
Google releases Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber models with delayed 3.5 Pro Model release
- Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
- Google Releases Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber: A Cheaper, More Token-Efficient Flash Tier Built for Agentic Workloads
- Google releases three new Gemini models — but no 3.5 Pro
- Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
- Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
- Google ships 3 new Gemini models. Just not the one everyone’s waiting for.
Relationships
Products & technology
- Google develops this model · 6 sources