Yacine: SWE benchmarks are the only ones people care about — total CS victory
yacineMTB · x · 2026-09-29
Yacine MTB observes that the only benchmarks anyone in AI seems to care about are software engineering benchmarks; everything else barely matters, which he calls a total victory for computer science — coding ability has become the single yardstick for model progress.
More from Models
- Heavy user: two days of near-constant Opus 5.5 use burned only 10% of weekly quota — CtrlAltDwayne · 2026-09-29
- Leaked OpenAI 'dot' details show raising phone to ear triggers ChatGPT Voice — koltregaskes · 2026-09-29
- Google to replace Gemini Gems with Skills starting November 17 — mark_k · 2026-09-29
- Carla v0.1.0: a local llama.cpp loom TUI for growing AI characters — max_paperclips · 2026-09-29
- Leaked OpenAI DevDay reveal called 'just a Grok bot / Meta Muse rip-off' — gaganghotra_ · 2026-09-29
- Running Jev at high frame rate with full-state snap inferences makes it a true System 1 — mathemagic1an · 2026-09-29