SPADE: Auto-generating progressively harder training environments for self-improving agents

_AndrewZhao · x · 2026-08-21

Introduces the SPADE framework, where a single model acts as both the Environment Designer and the Reasoning Agent. It writes executable, agentic environments that increase in difficulty as the agent improves, enabling automatic environment scaling for continuous self-improvement.

Related event: SPADE: Single-Model Self-Play Auto-Generates Training Environments for Continuous Self-Improvement(11 posts)→

Original post →

More from Research

Research channel →