NYU's DAGent: evaluate-then-grow planning lifts deep research agents by up to 5.8 points

newyorkuniversity · hf · 2026-10-02

NYU researchers introduce DAGent, a DAG-based multi-agent framework for deep research built on Evaluate-then-Grow incremental planning: an Orchestrator grows the task graph batch by batch, conditioning each expansion on confidence and uncertainty signals from completed nodes — unlike brittle Plan-then-Patch systems that commit hardest when evidence is weakest and waste compute on branches that shouldn't have been planned.

Key techniques:

Results: On BrowseComp-Plus, GAIA and xbench-DeepSearch, DAGent beats the strongest open-source baseline by 5.3/5.8/2.0 points at Qwen3-235B-A22B scale, replicating across four open backbones and extending to GPT-5 at 327K context. At 8B scale, DAGRPO adds 3.0 average Pass@1 over same-budget outcome-only GRPO. Same-architecture comparisons show evidence-conditioned planning reaches higher accuracy at lower token, tool-call and step footprints. Code is open-sourced.

Original post →

More from coding & agent

coding & agent channel →