How AA Benchmarks 26 Search API Products Across 13 Providers

ArtificialAnlys · x · 2026-10-06

Artificial Analysis shared links for its search evaluation: the leaderboard, methodology page (including how OpenAI Web Search is run), and Strup, the open-source agent harness used for all Search API rows. Search API providers can request inclusion.

The benchmark covers 26 Search API products across 13 providers, evaluating underlying web indexes, result presentation, and cost/speed/quality tradeoffs for agentic use cases like deep research and coding. Scores are the equal-weighted mean of DeepSearchQA, BrowseComp and AA-Omniscience (0–100), all with the same candidate answer model. Data as of Sep 28, 2026.

Related event: OpenAI Web Search Debuts at No.5 in Search Benchmark, ~$0.05 per Task(4 posts)→

Original post →

More from Models

Models channel →