LLM Museum: historical benchmarks and infographics of frontier models since GPT-4
delusion54 · reddit · 2026-09-08
A Reddit user built an AI-generated "LLM Museum" site collecting historical benchmark comparisons and infographics of frontier models from OpenAI, Anthropic, and Google since GPT-4, as a non-commercial exploration of the race's evolution.
More from Models
- No, the OpenAI agent didn't 'escape' its sandbox — it just messaged HuggingFace servers — danbri · 2026-09-08
- Gary Marcus says LLMs still haven't produced new linguistic generalizations, echoing Chomsky — GaryMarcus · 2026-09-08
- $10/mo buys ~37,800 DeepSeek V4 Flash calls; open models near Sol-level math by year-end? — teortaxesTex · 2026-09-08
- Hobby Blender benchmark: GPT-6-Astra one-shot scenes strikingly outperform other tested models — Gruku · 2026-09-08
- Same prompt showdown: ChatGPT vs Google Stitch vs Figma Make for UI generation — Tegadesigns · 2026-09-08
- GPT-6 Astra builds a full interactive 3D ankle atlas in one session — DeryaTR_ · 2026-09-08