SOURCE // NEWS

Google Launches Three New Gemini Models for AI Agents But 3.5 Pro Remains Missing

Google Launches Three New Gemini Models for AI Agents But 3.5 Pro Remains Missing

On Tuesday, Google DeepMind officially released Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber. #Gemini 3.6 Flash is positioned as Google's 'workhorse model,' promising enhanced capabilities in coding, knowledge work, and multimodal tasks while reducing token consumption by up to 17%, making it significantly cheaper than its predecessor, 3.5 Flash.

Among the other releases, Gemini 3.5 Flash-Lite stands as the most cost-effective option in its class, while Gemini 3.5 Flash Cyber is a specialized model fine-tuned for detecting and patching cybersecurity vulnerabilities. This cybersecurity-focused model will be exclusively available to governments and trusted partners under a limited access pilot program.

Google emphasized that the primary focus of these releases is to deliver the efficiency, low latency, and reliability required by enterprise customers building AI agents at scale.

The launch is notable not just for what Google shipped, but also for what it left behind. The update conspicuously lacks the long-anticipated upgrade to Google's flagship model, Gemini 3.5 Pro, which was last updated back in February.

Since then, rivals have moved at a breakneck pace. OpenAI has launched GPT-5.5 and begun rolling out GPT-5.6, while Anthropic has introduced Claude Opus 4.8, Claude Sonnet 5, and expanded access to its frontier Fable 5 model, highlighting the intense pressure on Google's release cycle.

Google had previously teased the Pro release during the 3.5 Flash launch in May, stating it was 'already being used internally' and expected to roll out 'next month.' However, recent reports from Bloomberg revealed that Google faced internal delays with 3.5 Pro as the model struggled to meet internal performance benchmarks.

Generally, Gemini Pro models represent Google's peak capabilities for complex reasoning and coding, whereas Flash models prioritize lower costs and faster response times for production-grade software.

Google DeepMind product lead Logan Kilpatrick mentioned on Tuesday that the company is actively testing Gemini 3.5 Pro with select partners and hopes to launch it 'soon.' He also revealed that the team has initiated its most ambitious pre-training run yet for Gemini 4.

[AgentUpdate Depth Analysis] While competitors like Anthropic and OpenAI continue to dominate the headlines with high-end reasoning capabilities, Google is strategically securing the developer pipeline by optimizing the cost-to-performance ratio at the edge. For the AI Agent ecosystem, this release highlights a major shift: enterprise adoption is moving away from raw, expensive intelligence toward production-grade cost-efficiency and low latency. Gemini 3.6 Flash's 17% token reduction directly addresses the financial overhead of running complex, multi-turn agentic loops. The delay of Gemini 3.5 Pro suggests that frontier reasoning models are hitting a temporary plateau, making specialized, ultra-fast, and vertically-optimized 'Flash' models the true engines of immediate commercial agent deployment.