AgentWebBench 评测多智能体协作

XiongChenyan · x · 2026-07-11

一篇 ICML 2026 论文介绍了 AgentWebBench,作者称其是首个用于评估 multi-agent coordination 的 benchmark,场景设定在 Agentic Web。

论文设定的任务规模包括:

原文链接 →

「研究」频道最新

更多「研究」频道 AI 资讯 →