Gemini 3.5 Flash Lite hits 89% vs GPT-5.6's 93% at 1/10 the cost

mrdbourke · x · 2026-08-19

In a custom eval task, @mrdbourke found gemini-3.5-flash-lite scored 89% versus GPT-5.6 Sol (max) at 93% — a small gap despite being 10x cheaper and 20x faster. He suggests trying Optima to run the comparison on your own custom evals.

Original post →

More from Models

Models channel →