LoHoSearch turns a 7.62M-entity knowledge graph into a harder benchmark for search agents

美团技术团队 · wechat · 2026-07-23

Meituan's LoHoSearch uses a 7.62M-entity knowledge graph to generate harder search tasks

Meituan's LongCat team released LoHoSearch, a new benchmark for search agents that automates question generation from a large Wikipedia knowledge graph instead of relying on human-designed prompts.

Performance is much lower than on BrowseComp:

Additional findings:

The team argues LoHoSearch is a better stress test for long-horizon search and context management, and the benchmark is open sourced.

Original post →

More from coding & agent

coding & agent channel →