Kimi K3 Hits 88/98 in Benchmarks; Qwen3.8-Max Outperforms Its Score
Correct_Tomato1871 · reddit · 2026-08-09
A Reddit user shared recent LLM benchmark notes comparing Kimi K3, Qwen3.8-Max, and Gemini 3.6 Flash.
Key takeaways: Kimi K3 achieved a high score of 88/98. Qwen3.8-Max is notably stronger than its raw score suggests. Meanwhile, Gemini 3.6 Flash appears to have regressed in performance compared to its 3.5 predecessor.
More from Models
- Tencent Releases UI-Mate-27B, a Desktop GUI Agent Model — tencent · 2026-08-24
- Sakana AI translation outperforms Google and DeepL in Japanese-English benchmarks — SakanaAILabs · 2026-08-24
- Developer haider makes his own LLM tier list after disagreeing with theo's rankings — haider1 · 2026-08-24
- Mystery OxAlpha Beats Claude; Alibaba Raises $10B for AI — 创业邦 · 2026-08-24
- OpenAI and Google cut LLM prices; mystery OxAlpha model beats Claude on DeepSWE — 创业邦 · 2026-08-24
- AI News Digest: DeepSeek Weekend Discounts, GPT-5.6 Sol Price Cut, Alibaba's $10B AI Raise — APPSO · 2026-08-24