GPT-6 Astra Autonomously Trains a Pen-Spinning RL Policy in 36 Hours
JasonMa2020 · x · 2026-09-17
Walter Zhu's fourth GPT-6 Astra test: a single prompt asking for pen spinning with a dexterous Sharpa hand, RL training in Isaac Lab, a self-created pen mesh, and a visualization video — with freedom to search the web and download papers. After running autonomously for a day and a half (including policy training), it produced a working pen-spinning RL policy and demo video, prompting Jason Ma to note Eureka was ahead of its time.
More from Embodied
- Reka open-sources RekaDaily-10k: 10,000+ hours of real egocentric household robot training data — artetxem · 2026-09-17
- Grounded API launches with SOTA hand-tracking (<1cm) and SLAM benchmarks — databoydg · 2026-09-17
- Robotics team turns an actuator race condition bug into the feature they needed — eigenron · 2026-09-17
- ArmSoM Sige 7 unboxed: an RK3588 board for edge AI and robotics — chrismatthieu · 2026-09-17
- Closed-room Hong Kong robotics workshop tackles foundation models, data engines, deployment — paigeinsf · 2026-09-17
- Innate OS open-sourced: an agentic OS for general-purpose robots under $1k — ycombinator · 2026-09-17