
Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
Quick Answer
Google DeepMind has launched the Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber models, enhancing AI agent performance with 17% lower token usage and improved efficiency.
Quick Take
The 3.6 Flash model excels in coding and multimodal tasks, while 3.5 Flash-Lite achieves 350 output tokens per second, making it ideal for high-throughput applications.
Key Points
- 3.6 Flash reduces output token usage by 17% compared to 3.5 Flash.
- 3.5 Flash-Lite achieves 350 output tokens per second, enhancing agentic workflows.
- 3.6 Flash offers lower costs at $1.50/1M input tokens and $7.50/1M output tokens.
- Enhanced safety features in 3.6 Flash improve resistance to jailbreaks.
- 3.5 Flash-Lite significantly outperforms 3.1 Flash-Lite in agentic tasks.
Source Excerpt
We’re introducing new Gemini models, including Gemini 3. 6 Flash, 3. 5 Flash-Lite and 3. 5 Flash Cyber.
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from Google DeepMind
See more →
Introducing Gemma 4 12B: a unified, encoder-free
Google DeepMind has introduced Gemma 4 12B, a unified, encoder-free multimodal model designed to enhance performance across various tasks. This model aims to streamline processes in AI applications by eliminating the need for traditional encoders, potentially improving efficiency and reducing costs for developers and researchers in the field.
