Tavily's 93% live-retrieval agent study challenged: prompts were domain-anchored

edwin · x · 2026-10-09

Tavily's Agent Visibility Report ran 38,000 journeys across 1,056 real business sites with agents like Claude Code and ChatGPT, finding 93% of answer claims are grounded in live lookups rather than training knowledge; sites readable by agents get recommended 2.6x more; blocking agents doubles web search's share of answers (12%→25%), and those answers are 3.7x more likely to miss every requested fact.

edwin pushes back on methodology: every test prompt was domain-anchored (e.g., "what subscription options are available at {domain}"), which naturally primes agents to fetch that site — so 93% is unsurprising. In their own data across millions of journeys, asking the same questions without the domain makes training data weigh far more heavily.

Original post →

More from Research

Research channel →