Agent Training Data: More Tools, Harder to Generate

Shahules786 · x · 2026-07-14

This thread discusses how to generate post-training data that genuinely enhances agent capabilities. The core insight is that in enterprise-level, multi-turn long tasks, the number of tools expands exponentially. A single domain might have 100+ tools, and three domains easily exceed 300. Simply stuffing the tool list into the context degrades task performance.

The thread highlights three papers:

The author notes these papers are already influencing their internal practices: synthesizing tasks from tool graphs and combining multiple models of varying capabilities to assess trainability, with the ultimate goal of integrating these methods into a customizable, automated data pipeline.

Related event: PlanBench-XL Tackles Agent Tool Retrieval at Scale(3 posts)→

Original post →

More from coding & agent

coding & agent channel →