Google has released three new models in its Flash tier—Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber—tuned for speed, cost efficiency, and high-volume agentic work rather than maximum reasoning depth.

  • Gemini 3.6 Flash uses 17% fewer output tokens than 3.5 Flash (up to 65% reduction on DeepSWE) and lowers the output price from $9.00 to $7.50 per 1M tokens, while scoring higher on DeepSWE (49% vs 37%) and MLE Bench (63.9% vs 49.7%).
  • Gemini 3.5 Flash-Lite delivers low-latency performance at 350 output tokens per second for $0.30/$2.50 per 1M tokens, beating the older 3 Flash on SWE-Bench Pro (54.2% vs 49.6%) and OSWorld-Verified (74.0% vs 65.1%).
  • Gemini 3.5 Flash Cyber is a specialized model for finding and patching software vulnerabilities, used in the CodeMender agent to find 55 unique V8 issues compared to 47 for mainline 3.5 Flash.

These models are available via the Gemini API, Google AI Studio, Android Studio, and GitHub Copilot, with 3.5 Flash Cyber currently limited to governments and trusted partners due to dual-use risks.