Amazon AGI's AutoGym auto-generates tasks, environments, and verifiers for agent RL training

omarsar0 · x · 2026-09-29

A new Amazon AGI paper introduces AutoGym, a framework that auto-generates complete RL gyms — task, executable environment, and verifier — from a small domain seed or past model trajectories.

Three key mechanisms:

The paper argues static task sets saturate and get contaminated, single-pass synthesis yields cosmetically hard tasks, and LLM judges are unreliable. Results span productivity and temporal-reasoning settings.

Original post →

More from Research

Research channel →