Grok Bot and Muse too dumb for business agents, says engineer comparing Claude Code

jdjohnson · x · 2026-09-28

JD Johnson says the models behind Grok Bot and Muse fail quickly when accuracy, completeness, and instruction-following are required. Cited user reports: Muse repeatedly asking for credentials it already has, forgetting recent deploy changes, ignoring 'tunnels don't work', and getting lazier. He argues Claude Code or Codex setups show a massive quality gap, and expects OpenAI and Anthropic to nail the format first.

Related event: Developer finds Grok Bot and Muse fall short of commercial agent standards(2 posts)→

Original post →

More from coding & agent

coding & agent channel →