Kimi K3 flunks a crossword task after 2 hours and $30 of prompting
Kurk_Lazaris · reddit · 2026-07-28
A Reddit user says Kimi K3 failed badly on a surprisingly practical task: building a crossword.
- They spent 2 hours and $30 trying to get the model to generate a crossword.
- The model reportedly misunderstood the algorithmic logic, the clue descriptions, and the grid construction constraints.
- It kept producing words that did not actually cross correctly.
- The poster asks whether this is simply too hard a task even for top-tier models, or whether the result reflects a broader weakness in Kimi K3.
More from Models
- Users discuss what they actually use Opus 5 for beyond coding — remilouf · 2026-07-29
- Anthropic Hints at Achieving Recursive Self-Improvement, Calls for Pacing Frontier — daniel_mac8 · 2026-07-29
- Claude Opus 5 tops DeepSWE with a 74% score and a claimed 28% cost edge — daniel_mac8 · 2026-07-29
- LiquidAI’s 230M LFM2.5 encoder trends on Hugging Face — LiquidAI · 2026-07-29
- Anthropic may be 1.5 generations ahead internally, with Fable 5.1 weeks away — haider1 · 2026-07-29
- Moonshot’s Kimi K3 is a 2.8T open-weight MoE model with 1M-token context — alex_verem · 2026-07-29