Jev's eval abstraction maps 1:1 to autorubric paper from 8 months ago, researcher finds
deliprao · x · 2026-09-21
Delip Rao found that typesafeai's Jev uses an evaluation abstraction strikingly similar to autorubric, a paper released 8 months ago and to be presented at COLM next month: Jev's Noul/Score/Choice criterion types map 1:1 to the paper's types. He frames it as "like minds think alike" and calls Jev an incredible model.
Related event: Researcher Alleges TypeSafe's Jev Mirrors His AutoRubric Paper(3 posts)→
More from Models
- Dev buys a Meta coding subscription for its 'excellent model, crazy quota, low price' — intellectronica · 2026-09-21
- Rumored Opus 5.2/5.5 outputs circulate; execs reportedly expect taste gap to close — teortaxesTex · 2026-09-21
- Astra for prose, Fable for long-horizon work: a developer's multi-model division of labor, with Pi as best harness — seatedro · 2026-09-21
- Zero-shot embedding classifiers: prototyping superpower or lazy black box? — antoine_chaffin · 2026-09-21
- Bespoke-Nimble-9B, a Qwen3.5-9B LoRA for evidence-grounded text classification, trends on Hugging Face — bespokelabs · 2026-09-21
- Sophia Yang Gets Jev Access, Runs Small RL Experiment With Fireworks AI — sophiamyang · 2026-09-21