Kimi K3 Beats Opus 4.8 in 34-Prompt Oneshot Eval at 1/16 the Cost

kms_dev · reddit · 2026-07-31

A developer ran 34 oneshot prompts through both Kimi K3 and Opus 4.8, evaluating the generated HTML, screenshots, and GIFs using Sonnet 4.6. The results showed Kimi K3 performing better than Opus 4.8.

Furthermore, Kimi K3 proved highly token-efficient, costing only $0.44 to process all 34 prompts compared to $7.16 spent by Opus.

Original post →

More from Models

Models channel →