Sentry founder on eval costs: keep asserts deterministic, LLM judges get brutal

zeeg · x · 2026-09-10

In an exchange with another developer, Sentry founder David Cramer (zeeg) shares how he keeps LLM eval costs down: most of his assertions are deterministic, rubrics are minimal, and only a few non-deterministic asserts require running an agent. The counterpart complains that testing 10 models across 3 evals with LLM-as-judge is already expensive, and worries costs will balloon 100x with more evals and pricier models — a candid look at how brutal agent eval iteration costs have become.

Related event: Sentry Founder Shares LLM Eval Cost Lessons(3 posts)→

Original post →

More from coding & agent

coding & agent channel →