Shipping AI Features: The Messy Reality Behind the 'Pick a Model and Launch' Fantasy

aakashgupta · x · 2026-10-11

Aakash Gupta contrasts the fantasy of shipping an AI feature (pick a model, write a prompt, build a demo, launch) with the messy reality: CEOs force demos, sales pre-sells, real users type with typos, outputs go generic without thread/CRM context, the model invents refund policies, legal gets involved, and a 95%-accurate eval judge catches zero actual failures. After shipping to 5%, usage stays flat, one power user eats the margin, and every new model release forces a full eval re-run.

His core point: what decides success sits around the model — context, tools, permissions, and checks that catch failures before customers do. Often the model was never the problem.

Related event: Shipping AI Features Is Messier Than You Think(2 posts)→

Original post →

More from coding & agent

coding & agent channel →