Kimi K3 Hits 295 tok/s in Tests, Confirmed Full Precision Without Quantization

JiaZhihao · x · 2026-08-10

Addressing community questions about Kimi K3's performance on the Artificial Analysis benchmark, Lithos AI confirmed the results are valid, reaching an impressive 295 tok/s.

The team emphasized that they did not use funky quantization to achieve this speed, maintaining native precision and full model quality. Additionally, the pricing is not based on unrealistic batch-size-1 assumptions. The public API is not fully available yet as they are currently ramping up capacity.

Original post →

More from Models

Models channel →