Kimi K3 Wildly Overestimates Project Workload
doodlestein · x · 2026-07-20
This post jokes about **Kimi K3**'s "thinking traces": the model estimated the author's **FrankenSim** project at **27,500 hours**, though the author's actual goal was just a month or two. The attached screenshot shows the model's thinking interface, listing the number of questions, estimated hours, dependencies, and some "authenticity/consistency" bug stats. Overall, it highlights the model's exaggerated judgment regarding project scale and timelines during reasoning.
Related event: Kimi K3 Thinking Traces Go Viral for Mockery and Wild Estimates(2 posts)→
More from Models
- Kimi K3 and Fable 5 show nearly identical failure patterns on a software benchmark — FinanceYF5 · 2026-07-21
- Kimi K3 hits 89.4% peak on software tasks while Fable 5 is slightly steadier — FinanceYF5 · 2026-07-21
- Kimi K3 leads on Go, but Fable 5 wins Python, JavaScript, TypeScript and Rust — FinanceYF5 · 2026-07-21
- Kimi K3 costs $4.65 per run and delivers 2.8× more work per dollar than Fable 5 — FinanceYF5 · 2026-07-21
- Kimi K3 reaches 89.4% pass@4 and tops the benchmark over GPT-5.6 Sol — FinanceYF5 · 2026-07-21
- Kimi K3 and Fable 5 now look much closer than the old open-vs-closed gap — FinanceYF5 · 2026-07-21