Practitioner: models ace verifiable tasks but sales-intel workflows hit ~5% accuracy

JayKurtz90 · x · 2026-09-14

A practitioner working across research, writing, sales and knowledge engineering argues that models keep improving on verifiable tasks but still lack domain-specific reliability in dynamic environments. Examples: "sales intelligence" platforms he tested achieve only 5% accuracy on what actually moves deals, and no product like Astra will magically generate research pre-processed for his own follow-up questions. He finds these custom workflows highly valuable but very hard to build, spent his day rebuilding his first attempt at a global business workflow, and asked for a dedicated lab/course on the topic.

Related event: Practitioner flags gap in domain reliability as models excel at verifiable tasks(2 posts)→

Original post →

More from coding & agent

coding & agent channel →