Agent runs autonomously for 1.5 days to train a dexterous-hand pen-spinning RL policy
SatelliteNetSec · x · 2026-09-18
In a fourth test of what's called GPT-6 Astra, walterzhu8 gave an agent a single prompt: implement pen spinning on a Sharpa dexterous hand using Isaac Lab for RL training, create the pen mesh, and freely search the web and download papers as needed. The agent ran autonomously for a day and a half — including training the policy itself — and delivered a trained RL policy plus a visualization video, with credit to student Chengyang Li. LiveOverflow's reshare quips that "pen testers will soon lose their job," underscoring how far long-horizon autonomous agents have come on real engineering tasks.
Related event: GPT-6 Astra Autonomously Trains Dexterous Hand to Spin a Pen in 36 Hours(3 posts)→
More from coding & agent
- Salesforce launches Trusted Enterprise AI Harness to unify agent context, governance and security — emmanuelvivier · 2026-09-18
- Running Codex, Claude and Pi Agents Safely: gVisor Sandboxes Plus tart macOS VMs — craigbalding · 2026-09-18
- EvalSeal: open-source tool shows LLM judges flip verdicts on 5 of 20 borderline eval cases — Fit_Fortune953 · 2026-09-18
- Armin Ronacher floats replacing MCP with codemode + OpenAPI + RAG over API docs — mitsuhiko · 2026-09-18
- Obsidian Starter Kit v4 ships with MCP server, osk-cli and ~375 specialized AI skills — dSebastien · 2026-09-18
- Retrying LLM Requests Isn't Always Safe: Gateway Policies for Partial Streams and Side Effects — Rama_Surasani_ · 2026-09-18