Google Launches Three Gemini Models at Once: 3.6 Flash, 3.5 Flash-Lite, 3.5 Flash Cyber
Summary: Google expanded its Gemini lineup on July 21 with three new models covering performance, budget, and security use cases simultaneously.
Key Facts
- Gemini 3.6 Flash: Up to 65% fewer output tokens than 3.5 Flash, DeepSWE coding score jumps 37→49, knowledge cutoff advanced to March 2026. Priced at $1.50/$7.50 per million tokens (input/output)
- Gemini 3.5 Flash-Lite: Cheapest tier at $0.30/$2.50 per million tokens — targeted at lightweight tasks and high-frequency API calls
- Gemini 3.5 Flash Cyber: Security-tuned model for finding and fixing vulnerabilities, restricted to governments and trusted partners in a limited pilot
- Both 3.6 Flash and 3.5 Flash-Lite are live now in Google AI Studio, Android Studio API, and GitHub Copilot
- Flagship Gemini 3.5 Pro was conspicuously absent again; Google hinted that Gemini 4 pre-training is already underway
Why It Matters
The 65% output-token reduction in 3.6 Flash could cut costs in half for agentic workflows that repeatedly call long contexts. Separately launching a security-hardened variant signals Google's intent to compete directly in the vertical AI security market — a move rivals have not yet matched at this scale.
Read More
- Google releases three new Gemini models — but no 3.5 Pro — TechCrunch
- Gemini 3.6 Flash cuts AI agent token costs by up to 65% — VentureBeat
- Official Google announcement — Google Blog