Dev calls for speed/productivity benchmarks: frontier LLMs still ship with no sense of time

gandamu_ml · x · 2026-09-17

gandamuml argues it's wild that frontier LLMs still ship with no sense of time. Instead of task-completion metrics and pelicans-on-bikes style benchmarks, he suggests scoring models on speed and productivity.

Original post →

More from Models

Models channel →