Building a gold-standard eval set with zero users: the day-zero dataset dilemma

Illustrious-Roll9476 · reddit · 2026-09-19

A prelaunch developer hits the classic eval cold-start problem: the standard advice is to build datasets from production logs, but with zero users the ground truth has to be manufactured from scratch.

Asks recent launchers: go full synthetic at day zero, or ship and fix in production?

Original post →

More from coding & agent

coding & agent channel →