Zhipu Open-Sources Slime RL Framework

Zhipu has open-sourced the Slime framework, a reinforcement learning training stack used for its GLM models. It introduces a deterministic train-rollout alignment path that resolves the long-standing numerical mismatch between training and inference in large model RL.

2026-08-11 ~ 2026-08-12 · 2 related posts