Google has launched its new Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber models, designed to improve token efficiency, reduce latency, and enhance performance for production AI agents.
Key Points
- Gemini 3.6 Flash reduces output token usage by 17% compared to 3.5 Flash while improving coding and multimodal performance.
- Gemini 3.5 Flash-Lite is the fastest model in the series, delivering 350 output tokens per second for high-throughput agentic workflows.
- Gemini 3.5 Flash Cyber is a specialized model integrated into the CodeMender platform to identify and patch cybersecurity vulnerabilities.
- 3.6 Flash and 3.5 Flash-Lite are available immediately via the Gemini API, Google AI Studio, and the Gemini Enterprise platform.
- The company has initiated pre-training for the next-generation Gemini 4 model and is currently testing Gemini 3.5 Pro with partners.