Robotics RL Tutorial: Training Agents to Balance Using PPO

ShawnHymel · x · 2026-08-06

Creator Shawn Hymel has released part 2 of his Reinforcement Learning for robotics series.

The video introduces the PPO (Proximal Policy Optimization) algorithm and provides a hands-on demonstration of how to use it to train an agent to balance a robot in a simulated environment.

Original post →

More from Embodied

Embodied channel →