CUHK Team Open Sources Libra: 3x Throughput for Agentic Training
jiqizhixin · x · 2026-08-21
CUHK team released Libra to fix resource contention in agentic RL post-training. It uses a global planner to dynamically shift GPUs between rollout and training via an elastic hybrid pool. It also uses a causality-driven scheduler based on tool-return signals. Tested on 48 A800 GPUs, Libra achieves 3.0x higher throughput and converges 2.5x faster in reward.
More from Infra
- Unsloth V3 Qwen Model Broken on Dual AMD GPUs, V2 Fix Available — Equivalent-Ear-8016 · 2026-08-21
- Open-source x402-cleanweb-agent saves 80% tokens by cleaning web content — EstablishmentTough18 · 2026-08-21
- Pretraining Potential: Coding Agents and the Compute Bottleneck — zeeshanp_ · 2026-08-21
- The Math: Claiming 100T Tokens/Day Would Need ~580K GPUs — teortaxesTex · 2026-08-21
- Moore's Law Fading: Non-Silicon Computing and Novel Architectures to See Capital Influx — MikePFrank · 2026-08-21
- "Why do we need more datacenters? Just write faster kernels" — basedjensen · 2026-08-21