Agent runs autonomously for 1.5 days to train a dexterous-hand pen-spinning RL policy

SatelliteNetSec · x · 2026-09-18

In a fourth test of what's called GPT-6 Astra, walterzhu8 gave an agent a single prompt: implement pen spinning on a Sharpa dexterous hand using Isaac Lab for RL training, create the pen mesh, and freely search the web and download papers as needed. The agent ran autonomously for a day and a half — including training the policy itself — and delivered a trained RL policy plus a visualization video, with credit to student Chengyang Li. LiveOverflow's reshare quips that "pen testers will soon lose their job," underscoring how far long-horizon autonomous agents have come on real engineering tasks.

Related event: GPT-6 Astra Autonomously Trains Dexterous Hand to Spin a Pen in 36 Hours(3 posts)→

Original post →

More from coding & agent

coding & agent channel →