Kimi Ditches Positional Encoding Entirely at Scale

tokenbender · x · 2026-07-28

Developers analyzing Kimi's large model architecture discovered that it completely removes Rotary Position Embedding (RoPE) at its current scale, eliminating this baked-in positional bias. The author views this as another victory for the 'Bitter Lesson' in AI scaling.

Related event: Kimi K3 Architecture Drops RoPE for NoPE(3 posts)→

Original post →

More from Models

Models channel →