Probability-only model tested as LLM judge: 25x faster, 600x cheaper than Fable-class

danshipper · x · 2026-09-16

Dan Shipper (Every) says his team spent a week testing a new foundation model that outputs probabilities instead of text, making it an efficient LLM judge — in their tests 25x faster and 600x cheaper than a Fable-level model. Commenters see it as a "perfect workflow model" for agent skills needing evaluation, with Shipper calling it indispensable within 6-12 months.

Related event: Every tests Jev, a probability-only foundation model 25x faster and ~600x cheaper than Fable(7 posts)→

Original post →

More from coding & agent

coding & agent channel →