Google launches Gemini 3.6 Flash and 3.5 Flash-Lite for production AI agents
GoogleAI · x · 2026-07-21
Google says it is introducing two new Gemini models to balance efficiency and quality for production AI agents:
- Gemini 3.6 Flash: improves coding, knowledge work, and multimodal tasks, while using substantially fewer tokens per task.
- Gemini 3.5 Flash-Lite: the fastest and most cost-effective 3.5-class model yet, tuned for agentic workflows and reaching about 350 output tokens/sec with improved coding and overall quality.
Google says developers can start using both models now through the Gemini API in Google AI Studio, and they are also available in the Gemini app.
Related event: Google Launches Gemini 3.6 Flash and Other New Models(59 posts)→
More from coding & agent
- Devin Outposts aims to run AI agents on any machine, from Mac minis to Kubernetes clusters — blaizedsouza · 2026-07-22
- Devin adds e2b sandboxes for remote agent execution — badphilosopher · 2026-07-22
- Hermes Agent Refactoring Proposal: Decoupling via Event Bus and Monorepo Slicing — Promptmethus · 2026-07-22
- ty now reads Pydantic config keywords and field metadata — charliermarsh · 2026-07-22
- Pensar Launches AI Security Agent to Autonomously Discover and Patch 0-Days — andriy_mulyar · 2026-07-22
- ty adds first-class Pydantic support, including strict and lax field handling — charliermarsh · 2026-07-22