OpenAI’s internal Codex also found the counterexample in one autonomous run
aaron_lou · x · 2026-07-21
A thread claims OpenAI’s internal Codex also found the same counterexample, and that the result was produced fully autonomously in a single shot.
- The model reportedly completed the work without any harness.
- The prompt has been publicly released.
- The update strengthens the case that the result was not a human-assisted artifact, but an actual autonomous model run.
More from coding & agent
- Dev builds talk on guardrails workflow for shipping AI-written code without reading it — TejasKumar_ · 2026-09-11
- banteg: Codex auto-review has regressed, blocking steps needed to complete authorized tasks — banteg · 2026-09-11
- lucasmeijer's workflow: handwrite the doc yourself, then have the agent challenge your understanding — lucasmeijer · 2026-09-11
- A doc-anchored agent workflow: you write, the agent only critiques and finds disagreements — lucasmeijer · 2026-09-11
- SymKit MCP: 44 tools for AI agents to verify symbolic derivations — Foreign-Specific-604 · 2026-09-11
- GitHub Copilot team routes user bug reports to an AI agent via Slack — marlene_zw · 2026-09-11