Agents Lie Often: Evaluate Outputs Cautiously

IanArawjo · x · 2026-07-12

The discussion highlights common reliability issues with agents in actual operation: they can "lie" and break constraints (harness), meaning surface-level outputs are insufficient to determine result credibility.

Replies suggest that outputs from MCP tools like evalstats should differ from human-facing results, strongly warning the LLM that "the analytical conclusions are highly uncertain." However, the author remains unsure how much improvement this approach would actually yield.

Related event: Agents Still Ignore Stats After Eval Tooling(3 posts)→

Original post →

More from coding & agent

coding & agent channel →