125M 'decision' models matched 200x larger LLMs — prior-art spat shadows Jev
zmkzmkz · x · 2026-09-25
After Delip Rao noted typesafeai's Jev shares its abstraction with autorubric, researcher ashabrawy201 pointed to their 2024 Statement Tuning paper: 125M universal 'decision' models matching 200x larger LLMs on unseen tasks, with inference mapping to Jev's decision head — a prior-art debate around the small-model-plus-decision-head paradigm.
More from Research
- Puro-2B: an open recipe trains a Qwen2-1.5B-beating LLM on RTX 5090s for just $4.4K — IgorCarron · 2026-09-25
- Oxford-led NeurIPS paper: frontier LRMs mirror human rule discovery and brain activity — sreejan_kumar · 2026-09-25
- bold_lab_ai hiring: postdocs and RAs due Oct 9, strategic partnership PM due Oct 14 — josephdviviano · 2026-09-25
- AIRA₂ Research Agents Hit 81.5% on MLE-bench-30, Beating Prior SoTA of 72.7% — mariofilhoml · 2026-09-25
- Overfitting the Validation Set Is Rarely a Big Issue, AIRA₂ Ablations Show — mariofilhoml · 2026-09-25
- Nature Machine Intelligence: Cybernetics and interoception framework for embodied AI agents — AnnaCiaunica · 2026-09-25