Google’s New Flash Model Breaks Its Own Price War

Google just released three new Gemini models on July 21, and the one that matters most is Gemini 3.6 Flash. It’s the workhorse model Google is betting the company will use for 80 percent of the work people actually do.

Here’s the thing: it got 17 percent faster while getting cheaper. Fewer output tokens. Lower price per token. Same or better capability. That’s the direction everyone else wants to move but can’t quite get there yet.

Priced at $1.50 per million input tokens and $7.50 per million output. Knowledge cutoff jumped from January 2025 to March 2026. Multimodal. Native video understanding. The model is built for what real users actually do, not what looks impressive in a benchmark.

Google also released Gemini 3.5 Flash-Lite, its most cost-effective option, and 3.5 Flash Cyber, a specialized model fine-tuned for finding and fixing cybersecurity vulnerabilities. The company is signaling it understands something important: the frontier is not one massive model. It’s a portfolio where you use the right tool for the right job.

Conspicuously absent: Gemini 3.5 Pro, Google’s flagship, which has now missed its target multiple times. That gap matters. Google is optimizing for efficiency and coverage instead of chasing raw capability. Might be the smarter move. The market will tell us.

——

Follow: @Ali Demi
Book your free AI clarity call, NOW!
https://buff.ly/TpWy277

——

Sources:
https://techcrunch.com/2026/07/21/google-releases-three-new-gemini-models-but-no-3-5-pro/
https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/
https://www.androidauthority.com/google-launches-gemini-36-flash-3689795/
https://9to5google.com/2026/07/21/gemini-3-6-flash-launch/

Repost this. Thanks.