Gemini 3.6 Flash Beats GPT-5.6 in Browser Agent Benchmark
allenainie · x · 2026-07-22
According to browseruse's benchmark tests, Google's newly released Gemini 3.6 Flash achieves a score of 68% in web agent tasks. It outperforms GPT-5.6-sol (67%) and Sonnet 4.6 (62%), ranking second only to Opus 4.8 (74%) at a fraction of the cost, offering frontier-level browser agents at Flash pricing.
Related event: Gemini 3.6 Flash Review: Faster and Cheaper, But Not Smarter(16 posts)→
More from coding & agent
- Inspired by OpenAI's 10,000-agent run, dev open-sources a crowdsourced agent problem-solving platform — Benjaminsen · 2026-09-11
- Lucid: open-source Mac app keeps your laptop awake only while AI agents run — Pitiful_Hedgehog_600 · 2026-09-11
- banteg's snail project crowdsources AI agents to finish matching Snail Mail's 20 remaining functions — banteg · 2026-09-11
- Alex Townsend posts 200 open problems in numerical linear algebra for humans and AI agents — IgorCarron · 2026-09-11
- Kimi K2.8 Preview rolls out: near-K3 coding performance, 1M context for all tiers — teortaxesTex · 2026-09-11
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11