How to eval an AI agent before you have users: synthesize realistic queries first

hugobowne · x · 2026-09-24

A practical playbook for cold-start agent evals: define who your users are and what they're doing, feed that context to a model to generate candidate queries, curate them by hand, and mine existing sources like support tickets and onboarding docs. Run them through your prototype, log passes/failures, and you've got a seed eval set for every prompt or tool change.

Original post →

More from coding & agent

coding & agent channel →