Kimi K3 Generates a 32-Page Chain of Thought
emollick · x · 2026-07-19
While testing Kimi K3, the author asked it to recommend two relevant poems about the "current state of GenAI." The model generated a chain of thought spanning 32 pages, which the author noted contained numerous loops and dead ends.
The key takeaway here isn't the poetry itself, but K3's reasoning performance on complex, open-ended tasks: the thinking process is incredibly lengthy, but its overall stability is somewhat lacking.
Related event: Testing Kimi K3: 32-Page CoT Loops and English-Dominant Reasoning(5 posts)→
More from Models
- Same Echo Maze prompt, three frontier models: all passed visually but shipped the same hidden bug — eyishazyer · 2026-09-11
- Benchmark scores drop from 89% to 19% on new evals — how benchmaxxing breaks leaderboard trust — airesearch12 · 2026-09-11
- ChatGPT tells user their question is too hard and to 'accept dumber answers' — phido3000 · 2026-09-11
- Claude is no longer available for minors as Anthropic rolls out age assurance — Muhammad523 · 2026-09-11
- Developer Building a Unified Leaderboard of All Model Benchmark Scores — airesearch12 · 2026-09-11
- Rumor claims Kimi faked performance by serving Claude; DeepSeek new model surprises in evals — realsohamparekh · 2026-09-11