Cua offers scaled computer-use agent fleets, but top frontier agent clears just 6 of 25 KiCad tasks

lucasmeijer · x · 2026-09-11

Lucas Meijer asked what to know before wiring computer/browser use into his agents, pointing to Cua: a platform for running computer-use agent training, evals, and data generation at scale across Linux, Windows, macOS, and Android VMs, with snapshot forking, failure reproduction, and fleet pools that scale to zero. Its Cua-Bench shows the best frontier agent clears only 6 of 25 expert KiCad tasks — a sobering signal on real-world computer-use reliability.

Original post →

More from coding & agent

coding & agent channel →