Kimi K3 Leads Multi-Sampling Benchmarks

zephyr_z9 · x · 2026-07-18

The image compares the performance of **Kimi K3** and **Fable 5** across different sampling attempts: - `pass@1`: Kimi K3 scores 68.5, slightly below Fable 5's 69.9. - `pass@2`: Kimi K3 scores 82.0, beating Fable 5's 80.2. - `pass@4`: Kimi K3 scores 89.4, ahead of Fable 5's 88.5. The post highlights that Kimi K3 performs significantly better under multi-sampling settings, earning it the title of "open/closed SOTA".

Related event: Moonshot's Kimi K3 Tops Frontend Code Arena, Nearing Fable 5 in Coding at a Third of the Cost(12 posts)→

Original post →

More from Models

Models channel →