Particle.news
Download on the App Store

Google Ships Three New Gemini Flash Models as Pro Flagship Remains in Tests

The releases aim to cut token costs for large-scale AI agents through faster, more efficient models.

Overview

  • Google announced Tuesday, July 21, 2026, the immediate availability of Gemini 3.6 Flash, Gemini 3.5 Flash‑Lite, and Gemini 3.5 Flash Cyber across developer and consumer channels.
  • Gemini 3.6 Flash reduces output token use by about 17% versus 3.5 Flash and is priced at $1.50 per million input tokens and $7.50 per million output tokens to lower inference costs for agentic workloads.
  • Gemini 3.5 Flash‑Lite targets high‑throughput, low‑latency tasks with roughly 350 output tokens per second and pricing at $0.30 per million input tokens and $2.50 per million output tokens.
  • Gemini 3.5 Flash Cyber is a security‑specialized model tuned to find and patch software flaws inside Google’s CodeMender framework and will be available only to governments and trusted partners in a limited pilot.
  • Google said Gemini 3.5 Pro remains in partner testing after internal benchmark shortfalls, and DeepMind has begun a large pre‑training run for Gemini 4 as the company leans on efficiency‑focused models while it refines its flagship offering.