Google DeepMind launches Gemini 3.6 Flash with cheaper tokens, plus 3.5 Flash-Lite and Cyber models for coding and security
By admin | Jul 21, 2026 | 2 min read
On Tuesday, Google DeepMind unveiled Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber. The Gemini 3.6 Flash serves as Google’s "workhorse model," delivering enhanced capabilities in coding, knowledge work, and multimodal performance while cutting token usage by up to 17%, making it more cost-effective than its predecessor, 3.5 Flash. The Gemini 3.5 Flash-Lite stands out as the most budget-friendly option in its class, while the 3.5 Flash Cyber is a specialized variant fine-tuned to detect and address cybersecurity vulnerabilities at an affordable price. According to Google, this model will be available exclusively to governments and trusted partners through a limited access pilot program.
Google emphasizes that these releases are designed to provide efficiency, low latency, and reliability for customers building AI agents at scale. What makes this launch notable is not just what Google shipped—cheaper, faster models optimized for coding, efficiency, and cybersecurity—but what it omitted. The update does not include the long-awaited upgrade to Google’s flagship model, Gemini Pro, which was last refreshed in February. Since that launch, competitors have accelerated their pace: OpenAI released GPT-5.5 and began rolling out GPT-5.6, while Anthropic launched Claude Opus 4.8, Claude Sonnet 5, and expanded access to its frontier Fable 5 model, highlighting the intense release cadence among rival labs.
Google had previously hinted at the Pro release during the 3.5 Flash launch in May, stating that the Pro version was "already being used internally, and we look forward to rolling it out next month." However, last week Bloomberg reported that Google was grappling with internal delays in launching 3.5 Pro due to struggles in meeting internal performance targets. Gemini Pro models are typically Google’s highest-capability offerings, designed for complex reasoning and coding tasks, while Flash models prioritize lower costs and faster response times for production applications.
On Tuesday, Google DeepMind product lead Logan Kilpatrick noted that the company is currently testing Gemini 3.5 Pro with partners and hopes to "land soon." He also mentioned that the team has embarked on its most ambitious pre-training run yet for Gemini 4.
Comments
Please log in to leave a comment.
No comments yet. Be the first to comment!