OpenBMB says MiniCPM-Robot jumps from 10 Hz to 36 Hz with CUDA Graphs

iamfakhrealam · x · 2026-07-24

OpenBMB says PhyAI's CUDA Graph optimization and custom Triton fused kernels lift MiniCPM-Robot inference from 10 Hz to 33 Hz, and to 36 Hz on an NVIDIA H20.

The post argues that the key issue in embodied AI is not just perception, but making multi-step, context-aware inference cheap enough to run at robot speed. It also emphasizes that the release is fully open source.

Related event: OpenBMB Open-Sources MiniCPM-Robot for Offline Embodied AI(10 posts)→

Original post →

More from Embodied

Embodied channel →