Sentry CEO David Cramer: don't mock LLM calls, exercise real models in test harnesses

zeeg · x · 2026-10-07

Sentry CEO David Cramer (zeeg) argues against mocking LLM calls in tests: "Are you going to mock every database call?" Mocking expected output shapes forces constant updates whenever the model or behavior routes change. He distinguishes this from qualitative evals — evals are just LLM calls judging results, but if you're building a harness you need to actually exercise the models you use.

Related event: Sentry CEO: AI Agent Test Suite Costs $10 Per Run, Sparking Mock-vs-Real-Model Debate(8 posts)→

Original post →

More from coding & agent

coding & agent channel →