Google’s Gemini 3.5 Flash-Lite pushes 350 tokens per second at $0.3 in and $2.5 out
eyishazyer · x · 2026-07-22
- Gemini 3.5 Flash-Lite is the cheap, fast tier.
- Google says it runs at 350 tokens/sec, with pricing at $0.3/1M input and $2.5/1M output tokens.
- Despite being the “lite” model, it beats the older 3 Flash on some benchmarks.
- The charts show gains on agentic terminal coding and knowledge work, plus better software-engineering and computer-use scores versus 3 Flash.
Related event: Google Unveils New Gemini Models Emphasizing Cost-Efficiency(118 posts)→
More from coding & agent
- Codex set up its own mailbox and negotiated with local vendors for a water softener — sebpaquet · 2026-07-22
- Dashboard screenshot shows 354 agents, 1.8B tokens, and an 80-day streak — jonathan_wilke · 2026-07-22
- Multica says multica-agent is now its top GitHub contributor — jiayuan_jy · 2026-07-22
- Open-source Agent Fieldbook turns Anthropic and OpenAI docs into reusable skills — whitewust · 2026-07-22
- Letta pitches OS-style long-term memory for local LLM agents — thisguyknowsai · 2026-07-22
- Agent builders should map task complexity before copying Claude Code-style harnesses — hugobowne · 2026-07-22