Dev's tests confirm OpenAI Codex computer and browser use fail a lot

jdjohnson · x · 2026-10-08

Developer jdjohnson ran the numbers on how often OpenAI Codex's computer use and browser use capabilities fail. His suspicion that failures felt frequent turned out to be true — the measured failure rate is indeed high, with data attached. A notable reliability signal for anyone relying on Codex for autonomous browser/computer tasks.

Original post →

More from coding & agent

coding & agent channel →