Gemini 3.5 Flash-Lite targets high-throughput agents at $0.30 input pricing
xiaohu · x · 2026-07-21
Google positions Gemini 3.5 Flash-Lite as its fastest, most cost-effective 3.5-class model.
- Built for high-throughput agent workloads such as large-scale product extraction, bulk document translation, and summarization
- Claims output speed of 350 tokens/sec
- Priced at $0.30 input / $2.50 output per million tokens
Related event: Google Launches Gemini 3.6 Flash and Other New Models(59 posts)→
More from Models
- Google launches three new Gemini models, including a cybersecurity system — Polymarket · 2026-07-22
- Google says information agents are coming to AI Pro and Ultra this summer — gaganghotra_ · 2026-07-22
- Poolside’s Laguna S 2.1 gets a two-week free run on Nous Portal — NousResearch · 2026-07-22
- Qwen3.8 Max Preview looks substantially better in a side-by-side test with Kimi K3 — curiousily_ · 2026-07-22
- Moonshot’s Kimi K3 reaches #5 on MathArena as the top open model — xeophon · 2026-07-22
- Google launches Gemini 3.5 Flash Cyber for CodeMender, with limited access for governments — GoogleAI · 2026-07-22