Jev passes 8/9 computer-use tasks, makes decisions 13.6x faster than Astra/Codex

iamrobotbear · x · 2026-09-17

francedot tested TypesafeAI's new Jev model for computer use with Cua Driver: it passed 8/9 executable tasks (vs 9/9 for Astra/Codex), with median decisions 13.6x faster and median task time 27% lower.

The author's takeaway: once a model has a usable representation of the UI, much of computer use becomes a general decision problem. Solid showing for a first public release, though the test was small.

Original post →

More from coding & agent

coding & agent channel →