Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
TLDR Dev ● Covered by 13 sources
Google released Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber models designed for efficient AI agent deployment. Gemini 3.6 Flash reduces output token usage by 17% compared to 3.5 Flash while costing $1.50 per million input tokens and $7.50 per million output tokens, with 3.5 Flash-Lite running at 350 output tokens per second at $0.30/$2.50 per million tokens. These models enable developers to build and scale agentic workflows at lower cost and latency.
Why it matters
New Gemini models 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber have been introduced. The 3.6 Flash model improves coding and multimodal performance while reducing output token usage and costs, while 3.5 Flash-Lite targets high-throughput tasks and 3.5 Flash Cyber is designed for cybersecurity applications.
Also covered by
- Google DeepMind — Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission
- Meta AI Blog — How Meta’s AI Models Are Powering the First Wave of Genesis Mission Projects
- MarkTechPost — Google Releases Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber: A Cheaper, More Token-Efficient Flash Tier Built for Agentic Workloads
- TechCrunch AI — Google releases three new Gemini models — but no 3.5 Pro
- Ars Technica — Google announces Gemini 3.6 Flash and cybersecurity AI, teases 3.5 Pro and Gemini 4
- Google DeepMind — Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
- The New Stack — Google ships 3 new Gemini models. Just not the one everyone’s waiting for.
- The Verge — Google launches a cheaper alternative to large AI security models like Mythos
- TLDR — Google is building a chip with Gemini baked into the silicon
- The Neuron — Google Building Gemini-Native Server Chip Called Frozen
- The New Stack — Google just bet its inference future on a chip built for one model
- TechCrunch AI — Google is working on a new AI chip designed to make Gemini more efficient