Dev builds a one-stop benchmark comparison page for small models (4B-190B)
nunodonato · reddit · 2026-09-05
Reddit user nunodonato, frustrated that small-model benchmarks are scattered across sources and often missing from Artificial Analysis, built a simple web page to compare them side by side with filtering.
- Covers models released since April 2024, sized 4B-190B parameters
- The tool is live at nunodonato.com/aibench and the author is soliciting feedback
More from Models
- From GPT-5.0 to GPT-6.0 Astra: one year of model progress, documented via the same game prompt — mt229 · 2026-09-05
- Users report ChatGPT's latest image model has gotten noticeably more restrictive — TipRich9929 · 2026-09-05
- Qwen launches Token Plan from $6/month with Qwen3.8-Max and all-modality access — JaynitMakwana · 2026-09-05
- Thread continues: asking Qwen3.8-Max-0902 to build and self-debug a tower defense game — JaynitMakwana · 2026-09-05
- Qwen3.8-Max-0902 tested: 2.4T params, 1M context, and it built three full apps — JaynitMakwana · 2026-09-05
- Codebase audit on GPT6 Astra high used just 41% of quota in 30+ minutes — Yamapama · 2026-09-05