Google Launches Gemini 3.6 Flash and Other New Models

Google has officially released three new models: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber. They are now available across Google AI Studio, Vertex API, and Gemini App. This update reflects adjustments made based on developer feedback, aiming to better support large-scale agent applications in real-world scenarios.

Key Details & Features

Positioned by the company as one of its most powerful models to date, Gemini 3.6 Flash is designed to deliver higher-quality results using fewer tokens at the same cost. Compared to the previous 3.5 Flash generation, the new model comes at a lower price ($1.50 per million input tokens and $7 per million output tokens) while fixing critical issues and reducing unnecessary tool calls. Meanwhile, 3.5 Flash-Lite stands out as one of the smallest and fastest models available, boasting an output speed approaching 350 tokens per second at the same price point as the soon-to-be-retired 2.5 Flash.

Benchmarks & Reactions

In terms of benchmark testing, comparison charts from organizations like Artificial Analysis and various developers reveal that Gemini 3.6 Flash performs exceptionally well across multiple agent evaluations, leading in overall intelligence. @bindureddy even dubbed it the "best chat model in the world." Additionally, Elon Musk engaged with the announcement on Twitter. Despite the high praise, some Reddit users noted that the actual performance of these new models seems to fall slightly short of Google's official marketing claims.

2026-07-21 ~ 2026-07-22 · 64 related posts

Full story(6 episodes)→

7 near-duplicate retellings: Acceptable-Debt-294 · BLCNYY · OwariDa · Keano5567 · osanseviero · rseroter · shyamalanadkat