Particle.news
Download on the App Store

Google Expands Gemini Flash Line With 3.6 Workhorse, Lite and Restricted Cyber Model

The company is pushing token efficiency and lower operating costs for production AI agents to make large-scale deployments cheaper and faster.

Overview

  • Google released Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber on Tuesday, July 21, and made 3.6 Flash and Flash-Lite immediately available through the Gemini API, Google AI Studio, Android Studio, Gemini Enterprise, and the Gemini app.
  • Gemini 3.6 Flash is positioned as the new workhorse that uses about 17% fewer output tokens than the prior Flash build and is priced at $1.50 per million input tokens and $7.50 per million output tokens to lower per-task costs.
  • Gemini 3.5 Flash-Lite targets high-throughput, low-cost workloads by producing roughly 350 output tokens per second and costing $0.30 per million input tokens and $2.50 per million output tokens for bulk document processing and agentic search.
  • Gemini 3.5 Flash Cyber is a vulnerability-finding and patching model tuned for CodeMender and will be restricted to governments and vetted partners in a limited pilot because of dual-use risks.
  • Google said Gemini 3.5 Pro remains in partner testing with no firm release date and confirmed it has begun pre-training for Gemini 4, a next-generation run that signals the company is shifting resources toward future flagship capability.