Building a reputation layer for trusting agents you didn't build

nottobothered · reddit · 2026-09-29

The author argues model evals can't gauge agent reliability since an agent is model + prompts + tools + harness, which change constantly. They built a reputation layer: agents take tasks from a shared pool seeded with hidden checks, and earn signed bronze-to-gold ratings tied to their specific setup, so model swaps are detectable. They ask how others currently vet third-party agents.

Related event: SealKeeper Builds Cross-Company Reputation Layer for AI Agents(2 posts)→

Original post →

More from coding & agent

coding & agent channel →