
Google CEO
Google Doubles Down on Cheaper, Faster AI as Its Flagship Model Stays Stuck in Testing
Google just made its clearest statement yet about where it thinks the AI market is actually heading, and it isn't toward bigger, more expensive models. The company announced three new Gemini models on Tuesday, all built around speed and cost efficiency rather than raw capability, while confirming that Gemini 3.5 Pro, the flagship model originally promised for June, is still in testing, according to SiliconANGLE's reporting on the launch.
Google describes Gemini 3.6 Flash as its new "workhorse" model, better than its predecessor at tasks like coding while reducing token usage, the units of text an AI model processes, by up to 17%. Gemini 3.5 Flash-Lite is positioned as the fastest and most cost-effective version of the 3.5 family yet, built specifically for high-volume workloads inside larger AI agent systems.
A Direct Shot at Anthropic's Cybersecurity Lead
The most strategically interesting release is Gemini 3.5 Flash Cyber, designed to detect and patch software vulnerabilities at a performance level comparable to rivals but at a lower price per token than larger models, according to AOL's coverage of the announcement. Tulsee Doshi, senior director of product management in Google's Gemini group, confirmed the model will initially be available exclusively to governments and trusted partners.
That positioning matters because Anthropic and OpenAI have both rolled out their own cybersecurity-focused models in recent months, part of a broader race among frontier labs to own the security use case. Google's response, once again, is to compete on price rather than match capability head-on, a pattern worth understanding alongside our what is Anthropic guide on how the major labs are differentiating their enterprise offerings.
Why Efficiency Is Suddenly the Whole Strategy
Google's timing is notable. The three new models arrived one day before Alphabet's earnings report, as the company continues facing pressure over the delayed Gemini 3.5 Pro launch and mounting competition from both American rivals and increasingly capable Chinese open-weight models like Kimi K3, which we covered in our piece on Kimi K3's subscription halt this week.
According to Artificial Analysis data cited in Google's own materials, Gemini Flash already undercuts comparable models from Anthropic, OpenAI, and Chinese rivals on cost, a positioning strategy that reflects a broader shift in how businesses are approaching AI spending. As one Google executive noted, companies are increasingly realizing that using more expensive tokens doesn't always produce better outcomes, a philosophy that echoes what we've covered in our AI pricing guide on matching model cost to actual task complexity.
Why This Matters for Business
I've advised companies on AI adoption for four years, and Google's strategy shift is worth taking seriously even if you're not a Google customer. When the industry's most well-resourced AI lab is explicitly betting that cost-efficiency, not raw capability, is what will win the next phase of enterprise adoption, that's a strong signal about where AI vendor competition is actually headed.
For businesses currently paying premium prices for flagship AI models on every task, Google's new lineup is worth testing specifically for high-volume, lower-complexity workflows where a cheaper, faster model may deliver comparable results at a fraction of the cost.
The Fast Version
Google released three new Gemini models built around cost and speed rather than raw power, while its delayed flagship model, Gemini 3.5 Pro, remains in testing. The new lineup includes a cybersecurity-focused model aimed directly at offerings from Anthropic and OpenAI, priced lower per token than larger models. The launch reflects Google's broader bet that efficiency, not capability alone, will define the next phase of enterprise AI competition.




