Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
Google has introduced three new additions to its Gemini model family: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber. These models are designed to provide the efficiency, low latency, and reliability required by developers and enterprises building scalable production AI agents and managing high-volume workflows.
Gemini 3.6 Flash serves as a workhorse model, offering improved coding, knowledge work, and multimodal capabilities while reducing output token usage by 17 percent compared to the previous version. It is priced at $1.50 per million input tokens and $7.50 per million output tokens, and it ships with enhanced frontier safety safeguards against chemical, biological, radiological, nuclear, and cyber offense misuses.
Gemini 3.5 Flash-Lite is the fastest and most cost-effective 3.5-class model, reaching 350 output tokens per second. Priced at $0.30 per million input tokens and $2.50 per million output tokens, it targets high-throughput tasks like agentic search and document processing. Additionally, Gemini 3.5 Flash Cyber is a specialized model fine-tuned for detecting, validating, and patching cybersecurity vulnerabilities. Paired with the CodeMender agent, it will initially be available through a limited-access pilot program for governments and trusted partners.
Gemini 3.6 Flash and 3.5 Flash-Lite are available immediately for developers through the Gemini API in Google AI Studio and Android Studio, for enterprises via the Gemini Enterprise Agent Platform, and for general users in the Gemini app. Meanwhile, Gemini 3.5 Pro is currently being tested with partners ahead of a broader release, and pre-training has begun for the next-generation Gemini 4.