TogetherAI Open-Sources XoRL for Zero Train-Inference Mismatch in MoE RL

TogetherAI open-sourced XoRL, a distributed RL framework that achieves zero train-inference mismatch for large MoE models by tightly aligning computation at the kernel level. It significantly improved Wordle task performance with Qwen3.6-35B-A3B, with River API outperforming Tinker.

2026-08-18 ~ 2026-08-18 · 4 related posts