AUTO-UPDATED

Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

Google has launched its new Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber models, designed to improve token efficiency, reduce latency, and enhance performance for production AI agents.

Key Points

  • Gemini 3.6 Flash reduces output token usage by 17% compared to 3.5 Flash while improving coding and multimodal performance.
  • Gemini 3.5 Flash-Lite is the fastest model in the series, delivering 350 output tokens per second for high-throughput agentic workflows.
  • Gemini 3.5 Flash Cyber is a specialized model integrated into the CodeMender platform to identify and patch cybersecurity vulnerabilities.
  • 3.6 Flash and 3.5 Flash-Lite are available immediately via the Gemini API, Google AI Studio, and the Gemini Enterprise platform.
  • The company has initiated pre-training for the next-generation Gemini 4 model and is currently testing Gemini 3.5 Pro with partners.

Why it Matters

These model updates provide developers with more cost-effective and efficient tools to scale complex, agentic AI workflows in production environments. By balancing lower latency with higher accuracy, these releases help businesses integrate advanced automation into tasks ranging from software security to large-scale data analysis.
Blog.google Published by Tulsee DoshiSenior DirectorProduct Management, on behalf of the Gemini team
Read original