Gemini 3.7 Flash tops ARC-AGI benchmark at $0.12 per task
rakyll · x · 2026-08-21
Google's new Gemini 3.7 Flash model achieved high scores on the ARC-AGI benchmark: 95.5% on ARC-AGI-1 ($0.12/task) and 84.6% on ARC-AGI-2 ($0.25/task). It stands out for its low cost relative to other frontier models while maintaining high performance.
Related event: Gemini 3.7 Flash Tops ARC-AGI at $0.12 Per Task(4 posts)→
More from Models
- DeepSeek Vision Model Spotted in API Directory, Not Yet Active — teortaxesTex · 2026-08-21
- Ox Alpha open-source model benchmarks near Anthropic level — liminal_bardo · 2026-08-21
- Benchmark claims GLM 5.3 matches frontier models at 1/3 the cost — zainhas · 2026-08-21
- User Feedback: GPT5.6 Lags Behind Fable in Strategy Tasks — henloitsjoyce · 2026-08-21
- Speculation suggests low TTFT due to smaller distilled model architecture — teortaxesTex · 2026-08-21
- o1 excels at fixing vision-grounded bugs in rendering pipelines — teortaxesTex · 2026-08-21