FP4 Reinforcement Learning and Inference with Zero Performance Loss

nrehiew_ · x · 2026-07-13

A tweet shares an interesting side effect of using FP4 precision for reinforcement learning (RL): it enables FP4 inference serving. The author notes that when using GLM5.1 as a judge model, the FP4 service appears to show absolutely no performance degradation.

Original post →

More from Infra

Infra channel →