Hands-on with Instinct, Grok Bots and ChatGPT Work: personal agents compared
illscience · x · 2026-08-20
An in-depth look at Instinct, Grok Bots and ChatGPT Work, arguing all three distill the pattern OpenClaw established this year: persistent agents with cloud(ish) computers, browser access, cached credentials, recurring loops, plus a top-level orchestrator that reports and dispatches across threads.
Key takeaways:
- Browser use is now good enough for most web work (modulo CAPTCHAs/2FA); the key product variable is the agent's presumptuousness and resourcefulness in recovering from failure
- Instinct is the most aggressive: it reset a password on its own to finish a shopping task; iMessage-first with one continuous relationship, but a single long-running thread constrains ambitious work
- Grok Bots feels like an enterprise agent platform that also works for prosumers; multi-account support is thoughtful, though group-chat metaphor may not be intuitive
- ChatGPT Work is the least aggressive (credentials/payments caution looks like compliance, not tech limits) but has the best interface balance and full-duplex voice
- Cloud vs local roughly maps to knowledge work vs coding: knowledge work gains hugely from an always-on cloud VM
Related event: Hands-On Comparison: Instinct vs Grok Bots vs ChatGPT Work(2 posts)→
More from coding & agent
- Agents fail to reconsider strategy during post-training execution — omarsar0 · 2026-08-20
- Leaked System Prompt: Domestic Giant's Client Uses 3-Layer Memory & MCP Routing — vista8 · 2026-08-20
- Addy Osmani uses Agent loops to triage 90 daily PRs — rseroter · 2026-08-20
- One-man startup uses GPT-5.6 for cross-project agent collaboration — every · 2026-08-20
- ZenML releases Agent Trace Viewer to inspect coding agent trajectories locally — strickvl · 2026-08-20
- Claude Code + Opus 5 saturates ARC-AGI-3 via falsifiable predictions — scaling01 · 2026-08-20