Google DeepMind Unveils Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
Google/DeepMind
Google DeepMind
Google DeepMind announced three new AI models: Gemini 3.6 Flash, 3.5 Flash-Lite, and a specialized 3.5 Flash Cyber. These models aim to boost efficiency, reduce latency, and improve performance in agentic workflows. Gemini 3.6 Flash uses 17% fewer output tokens and is cheaper than its predecessor. 3.5 Flash-Lite is the fastest model in the 3.5 series, while 3.5 Flash Cyber is designed for cybersecurity.
Google DeepMind has introduced three new models in the Gemini family: 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber. Gemini 3.6 Flash is positioned as a workhorse, offering superior performance in coding, knowledge tasks, and multimodal applications. According to the Artificial Analysis Index, the model consumes 17% fewer output tokens compared to 3.5 Flash, and in some tests like Datacurve's DeepSWE, savings reach up to 65%. Pricing is set at $1.50 per million input tokens and $7.50 per million output tokens, which is lower than 3.5 Flash. The model also shows improvements in code accuracy (DeepSWE: 49% versus 37%), machine learning research tasks (MLE Bench: 63.9% versus 49.7%), and computer capabilities (OSWorld-Verified: 83.0% versus 78.4%). Enhanced safeguards against attacks in the areas of chemical, biological, radiological, and nuclear (CBRN) threats and cyber intrusions have been added to the model. Gemini 3.5 Flash-Lite is the fastest model in the 3.5 series, achieving 350 output tokens per second. Its price is $0.30 per million input tokens and $2.50 per million output tokens. It surpasses 3.1 Flash-Lite in quality and shows significant gains in coding (Terminal-Bench 2.1: 54% versus 31%), long-context tasks (GDM-MRCR v2: 72.2% versus 60.1%), and real-world task execution (GDPval-AA v2: 1140 versus 642). The model also outperforms Gemini 3 Flash on several benchmarks, including SWE-Bench Pro (54.2% versus 49.6%) and OSWorld-Verified (74.0% versus 65.1%). Gemini 3.5 Flash Cyber is a highly specialized model for cybersecurity, built on 3.5 Flash and fine-tuned for finding and fixing vulnerabilities. It powers the CodeMender agent and achieves competitive performance on the CyberGym benchmark. The model will only be available to governments and trusted partners as part of a limited-access pilot program. Both models, 3.6 Flash and 3.5 Flash-Lite, are available starting today through the Gemini API, Google AI Studio, Android Studio, and Google Antigravity. Additionally, testing of Gemini 3.5 Pro is underway with partners, and pre-training for Gemini 4 has begun.
Source: Google DeepMind —
original
