Google Surprise-Drops Gemini 3.6 Flash with Cheaper Pricing for Agentic Tasks
AGI Hunt · wechat · 2026-07-21
Without official promotion, Google has quietly added two new models to AI Studio: Gemini 3.6 Flash and Gemini 3.5 Flash Lite.
- Gemini 3.6 Flash: Described as "our most intelligent model yet," it focuses on delivering consistent frontier performance in agentic and coding tasks. Priced at $1.50 input / $7.50 output per million tokens, it actually lowers the output cost compared to the existing Gemini 3.5 Flash ($9.00 output) while improving performance.
- Gemini 3.5 Flash Lite: Designed for large-scale, high-frequency agentic tasks and subagent workflows, emphasizing high throughput and low latency. Priced at just $0.30 input / $2.50 output, it is ideal for the heavy lifting of parallel tasks delegated by a main agent.
Additionally, the author noted conflicting descriptions for 3.6 Flash across different pages and a somewhat confusing version naming scheme. The Pro version is notably absent, while the two new models are already available for testing.
Related event: Google Launches Multiple Gemini Models for Enhanced Cost-Efficiency(67 posts)→
More from Models
- NVIDIA says Nemotron 3 Ultra scored 30/42 on the 2026 IMO problems — NVIDIAAI · 2026-07-22
- Gemma-4-26B-a4B reportedly beats Qwen3.6 and Qwen3.5 MoE fine-tunes — JLeonsarmiento · 2026-07-22
- OpenAI is reportedly briefing U.S. lawmakers on its next model family — kimmonismus · 2026-07-22
- Muse Spark 1.1 lands at 1495 on Text Arena with standout agentic-coding price performance — ycombinator · 2026-07-22
- Advanced AI Models Are Becoming Impossible to Plug and Play — emollick · 2026-07-22
- Google Gemini's AI Problem: No Leading Model for Core Workloads — bindureddy · 2026-07-22