Today’s launches are all about better performance, lower latency, and a smaller bill.
- 3.6 Flash cuts token usage by up to 65% on complex coding
- 3.5 Flash-Lite reaches speeds of 350 output tokens/sec
Both are live in the Gemini app today!
Next up: Gemini 3.5 Pro, which has officially entered partner testing.
Google DeepMind (@GoogleDeepMind)
We’re rolling out three new models to make AI agents faster, smarter, and cheaper at scale:
🔵 Gemini 3.6 Flash: It uses fewer tokens than 3.5 Flash to deliver higher quality work at the exact same cost.
🔵 Gemini 3.5 Flash-Lite: A fast, cost-effective option for everyday tasks like processing documents and agentic search.
🔵 Gemini 3.5 Flash Cyber: A cybersecurity model built to find and patch critical software vulnerabilities.
— https://nitter.net/GoogleDeepMind/status/2079589698490572961#m