Who Certifies AI Agents Before They Go to Work? The Case for an AI 'Bureau Veritas'

ThomasBuildLab · reddit · 2026-10-09

The author argues AI agents are still built like craftsman machines—hand-tuned prompts, ad-hoc testing, no professional qualification system—while increasingly handling finance, legal, infrastructure, and medical workflows judged only by benchmarks.

Core claim: being intelligent is not the same as being professionally qualified. The proposal: treat agents like aircraft going 'fit to fly', with competency exams based on human standards, simulated dangerous scenarios, supervised probation, independent safety audits, certification scoped to specific tasks and autonomy levels (invoice review ≠ payment execution), and recertification after any model/tool/config change—certifying the whole system, not just the LLM.

The author also predicts agents will shift from custom builds to standardized, Lego-like modules (reasoning, knowledge, memory, permissions, monitoring), but assembled systems still need end-to-end certification. The economic punchline: a new institution category—independent bodies whose product is institutional trust, backed by accumulated evidence of agent performance, recognized by employers, insurers, and regulators.

Original post →

More from coding & agent

coding & agent channel →