According to Google's announcement, the company released Gemini 3.6 Flash and Gemini 3.5 Flash-Lite today, with improved efficiency and lower costs than the previous generation. Gemini 3.6 Flash uses 17% fewer output tokens than 3.5 Flash, with pricing reduced from $9 to $7.50 per million output tokens, making it more cost-effective for running AI agents at scale.
However, Gemini 3.5 Pro, which Google promised at I/O 2026 in May for June delivery, remains in testing after falling short on internal coding benchmarks, per Bloomberg. Google confirmed it has begun pre-training for Gemini 4, described as "our most ambitious pre-training run yet." Both 3.6 Flash and 3.5 Flash-Lite are available today via the Gemini app, Google AI Studio, and API.