Researcher mocks time series benchmark culture: bechmaxxed models rarely match your domain
RexDouglass · x · 2026-09-19
RexDouglass snarks that time series and tabular prediction libraries share an unstated cultural faith: that a pile of benchmark-maxed models live in the same domain as your actual problem. He adds that it "only costs a quarter" to do something that would run on a TI-86 — a jab at heavyweight models being overkill for simple forecasting tasks. A pointed critique of benchmark culture's gap from real-world problems.
Related event: Researcher Mocks AI Hype and Benchmark Culture in Time-Series Prediction(2 posts)→
More from AGI Musings
- Ryan Greenblatt: Advanced AI Could Drive an Industrial Explosion With Output Doubling in About a Year — RyanGreenblatt · 2026-09-19
- Deep Dive: Open-weight models now within ~4 months of best closed frontier models — markjeffrey · 2026-09-19
- Greg Kamradt: Humans with AI still beat AI with AI on productivity — GregKamradt · 2026-09-19
- Coders embrace AI fully for code yet call AI-assisted writing a moral failure — AaronBergman18 · 2026-09-19
- Blogger predicts AI giants will wage war on drug discovery and aging by the 2030s — Dr_Singularity · 2026-09-19
- Musk predicts AI could push US GDP growth to ~4% next year, double current pace — Polymarket · 2026-09-19