Miles now supports RL training for Qwen and GLM models with high-performance kernels
ying11231 · x · 2026-08-28
Miles now supports RL training for Alibaba's Qwen3.8-Flash-Next and Zai's GLM-5.3-Flash. By pairing SGLang rollout with Megatron training, it provides SGLang-consistent, high-performance kernels. It implements QSA indexing and sparse attention for Qwen, and KDA + DSA with kpool-compressed indexing for GLM, validated end-to-end on GB300 GPUs.
More from Infra
- Australia Minister: No Fossil Fuel Carve-out for Datacenters — nordicinst · 2026-08-28
- ZED Camera priced at $500? DIY alternative costs just $150 — _William_F_ · 2026-08-28
- KOTOR Remaster Path Tracer Integrates DLSS 4.5 RR — Michael_Moroz_ · 2026-08-28
- NVIDIA's NVHBM Breaks the Die-Size Limit, 3-5x VRAM per GPU — Charuru · 2026-08-28
- Tutorial: Train a Raspberry Pi to Read Gas Meter Automatically with Neural Network — JeremyCMorgan · 2026-08-28
- Ninfer Benchmark: 5090 Doubles Throughput for Qwen3 27B — Rollingsound514 · 2026-08-28