Baseten Loops Adds Support for RL and Fine-tuning GLM-5.3-Flash

baseten · x · 2026-08-26

Baseten announced support for Reinforcement Learning (RL) and fine-tuning (SFT/OPD) of the GLM-5.3-Flash model on its Loops platform. The model is noted for being smaller than competitors like Kimi K3 and GLM 5.2, resulting in significantly cheaper inference. It has also been redesigned for inference efficiency at long context lengths, making it a strong candidate for task-specific RL.

Original post →

More from Models

Models channel →