Robot arm tests rank GPT-6 variants: better reasoning means better manipulation at 30x the cost
YuXiang_IRVL · x · 2026-09-23
In LIBERO manipulation tasks, GPT-6 Astra > Sol > Luna: stronger reasoning translates directly into better robot manipulation. But cost differs sharply — Astra spent $3.20 over 29 calls ($0.11/call) vs Luna's $0.23 over 62 calls ($0.0037/call), though Luna failed the task.
More from Embodied
- Tesla Berlin Confirms Workers Will Wear Camera Rigs to Train Optimus; Global Humanoid Shipments Top 22k in H1 — CyberRobooo · 2026-09-23
- eidon-ai releases tracker-pov robotics video dataset, 10K-100K entries — eidon-ai · 2026-09-23
- Dev turns a 90s TV into an AI TV with Raspberry Pi, mic, camera and GPT Realtime — OpenAIDevs · 2026-09-23
- Steering VLA robot models at inference time lifts grasp rate from 14% to 74% without retraining — drmapavone · 2026-09-23
- Frontier AI models can control robots to follow harmful requests, NBC News reports — ycombinator · 2026-09-23
- 'Robot-use' works on humanoids: coding agents zero-shot control a full humanoid for pick-and-place — ai · 2026-09-23