OpenAI web search can waste 87% of injected tokens, local pipeline matches 96% accuracy

Remote-Breadfruit204 · reddit · 2026-07-24

The author argues that OpenAI’s web search tool can be surprisingly expensive for agent loops because it injects a large amount of context back into the model.

They built a local, free alternative pipeline with:

On a subset of OpenAI’s SimpleQA benchmark, they report:

The code and eval report are available in the linked GitHub repo.

Related event: Open-source Webfetch slashes AI agent search costs and tokens(4 posts)→

Original post →

More from coding & agent

coding & agent channel →