Local web search tool says it cuts token use 87% and cost 66%

Remote-Breadfruit204 · reddit · 2026-07-23

A local web search stack claims 87% fewer tokens and 66% lower cost

The author built webfetch as an alternative to hosted search from Anthropic and OpenAI, arguing that the default model is expensive: $10 per 1k searches plus roughly 17k result tokens per query.

Key pieces of the system:

On its SimpleQA benchmark, the author says the same agent loop matched hosted search at 96% accuracy while using 87% fewer tokens and costing 66% less. One test loop of 16 searches reportedly saved $1.50. It ships as a PyPI install and can be added to Claude Code as an MCP server.

Related event: Local Webfetch Plugin Drastically Reduces Agent Search Costs(2 posts)→

Original post →

More from coding & agent

coding & agent channel →