- Gemini introduces three new Flash models: 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber in CodeMender, targeting higher token efficiency, lower latency, and reliable performance for AI agents.
- 3.6 Flash improves coding, knowledge work, and multimodal performance, reducing output token usage by 17% compared to 3.5 Flash, with lower cost per output token.
- 3.5 Flash-Lite is the fastest model in the 3.5 series, delivering 350 output tokens per second, priced competitively for high-throughput production tasks, and outperforming prior Flash-Lite generations.
- 3.5 Flash Cyber in CodeMender is a specialized cybersecurity model for detecting and fixing vulnerabilities, exclusively available to governments and trusted partners via a limited-access pilot.
- Gemini 3.5 Pro is in testing with partners, and the team has started pre-training for Gemini 4.
- 3.6 Flash shows performance gains in benchmarks like DeepSWE (49% vs. 37%), OSWorld-Verified (83.0% vs. 78.4%), and GDPval-AA v2 (1421 vs. 1349), with enhanced safety safeguards against misuse.
- 3.5 Flash-Lite outperforms 3 Flash in agentic and coding evals, such as SWE-Bench Pro (54.2% vs. 49.6%) and OSWorld-Verified (74.0% vs. 65.1%), and supports computer use as a built-in tool.
- Both 3.6 Flash and 3.5 Flash-Lite are available starting today via Gemini API, Google AI Studio, Android Studio, Gemini Enterprise, and the Gemini app.