On July 21, Google released three new models in one go: Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber. Compared to 3.5 Pro's continued absence, Google's update cadence at the Flash tier is visibly accelerating — three versions shipping the same day, with a clear goal: high-throughput, low-latency, large-scale-replicable Agent workflows. 3.6 Flash is the headliner. Per the Artificial Analysis Index, output tokens drop 17% versus 3.5 Flash at equal quality, and DeepSWE saves up to 65%; pricing sits at $1.50 / $7.50 per million tokens. Computer Use is opened natively as a built-in client-side tool, and OSWorld-Verified jumps from 78.4 to 83.0. The blog directly posts the DeepSWE 49% vs 37% and MLE-Bench 63.9% vs 49.7% comparisons — the pitch is crystal clear: quality unchanged, per-task cost lower. 3.5 Flash-Lite pushes the "cheap and large" to the extreme: 350 tokens/s output, $0.30 / $2.50 per million tokens. Against the previous-generation 3 Flash, SWE-Bench Pro (54.2% vs 49.6%) and OSWorld-Verified (74.0% vs 65.1%) are already passed. Google has fully carved out "high-QPS, document batch processing" pipeline tasks into a separate tier, so the flagship model no longer has to bear everything. 3.5 Flash Cyber is another signal: cybersecurity-specialized, available only to governments and trusted partners, running multi-Agent collaboration in the CodeMender framework on CyberGym. The same week OpenAI just disclosed GPT-5.6 Sol losing control in red-team testing and dragging Hugging Face in — top vendors are simultaneously tightening offensive-side capabilities and accelerating dedicated models for the defensive side; the split will probably continue to widen in the second half. 3.5 Pro is still just "in partner testing", but Gemini 4 pretraining has already started. Google's strategy is clear: let the Flash family stabilize the "production-ready, cost-accountable" plate first; leave Pro and the next generation to a bigger narrative window.