PufferLib 5.0 released: trains RL agents in under a second
jsuarez · x · 2026-09-14
Joseph Suarez releases PufferLib 5.0, the latest version of his open-source reinforcement learning library that can train super-human agents in under a second.
Related event: PufferLib 5.0 Released: Train RL Agents in One Second(4 posts)→
More from Research
- MIT, Stanford, Harvard and CMU researchers build Social Simulation Arena to benchmark simulators prospectively — _Hao_Zhu · 2026-09-15
- Post-training shift makes model distillation nearly impossible to detect, researcher argues — maksym_andr · 2026-09-15
- Insilico's AI-designed rentosertib shows biological age reversal across six proteomic clocks in Phase IIa — thione · 2026-09-15
- Anthropic's Claude formalizes Fermat's Last Theorem in Lean, largely autonomously in 11 days — thione · 2026-09-15
- DeepMind's AlphaGenome Atlas maps predicted effects of all 9B single-letter DNA variants, free — thione · 2026-09-15
- OpenAI says internal system solved 90-year-old Navier–Stokes Millennium Problem with a Lean proof — thione · 2026-09-15