Google Quietly Launches Gemini 3.6 Flash: Cheaper, Stronger, and Agentic-Focused
OwariDa · x · 2026-07-21
Without a formal announcement, Google has quietly added Gemini 3.6 Flash and Gemini 3.5 Flash Lite to AI Studio. Gemini 3.6 Flash is described as "our most intelligent model yet," focusing on agentic and coding tasks. It is priced at $1.50 input / $7.50 output per million tokens, making it cheaper than the previous 3.5 Flash's $9.00 output price, while offering stronger performance.
Meanwhile, Gemini 3.5 Flash Lite is designed for high-throughput, low-latency execution, targeting large-scale agentic tasks and subagent workflows. Priced significantly lower at $0.30 input / $2.50 output, it parallels the role of models like Claude 3.5 Haiku in handling heavy subagent lifting. Both models are currently accessible via AI Studio and Vertex API, though Google's naming conventions and documentation remain inconsistent, and the flagship 3.5 Pro is still missing.
Related event: Google Launches Gemini 3.6 Flash and Cost-Effective Models(59 posts)→
More from Models
- Google DeepMind launches Gemini 3.5 Flash Cyber for faster, cheaper code security — ralucaadapopa · 2026-07-22
- Poolside’s Laguna S 2.1 gets a two-week free run on Nous Portal — NousResearch · 2026-07-22
- Qwen3.8 Max Preview looks substantially better in a side-by-side test with Kimi K3 — curiousily_ · 2026-07-22
- Moonshot’s Kimi K3 reaches #5 on MathArena as the top open model — xeophon · 2026-07-22
- Google launches Gemini 3.5 Flash Cyber for CodeMender, with limited access for governments — GoogleAI · 2026-07-22
- Gemini 3.5 Flash-Lite beats 3.1 Flash-Lite on long-context retrieval in MRCRv2 — Dillonu · 2026-07-22