Decision model Jev makes web agents 4.3x faster and 13x cheaper across 240 runs
EdenEmarco177 · x · 2026-09-30
ora research tested Jev, a 150ms decision model by TypeSafe that answers closed choices at a fraction of LLM cost, against Claude Haiku 4.5 as the step decider for agents browsing a seeded site via browser-use, WebMCP and NLWeb.
- 10 tasks x 5 repeats across 3 protocols: 240 runs, each verified
- On average 3.6x faster (up to 4.3x) and 7.7x cheaper (up to 13x); faster in every paired run
- Task success held: 100% on WebMCP and NLWeb, browser 78% vs 71%
Takeaway: with deciding this cheap, the biggest gains come from what sites expose to agents. Full method and data are open-sourced.
More from coding & agent
- Redditor builds a standalone Windows game with ~150 Claude prompts — Ill-Range-4954 · 2026-09-30
- RemCTL 2.0 turns Apple Reminders into a native, open-source ChatGPT extension — rudrank · 2026-09-30
- SKATE: an open-source workshop memory OS built with Claude Code, plus a 3D-printed mic — lilweedbitch69 · 2026-09-30
- Claude Code offers $250 free cloud credits via /claim-credit command — daniel_mac8 · 2026-09-30
- AWS Shows How to Build a Multi-Agent Music Pipeline on Bedrock AgentCore Runtime Instances — AWS ML Blog · 2026-09-30
- Synapse: open-source agent memory with Markdown + SQLite, retrieval via MCP — BoringCelebration405 · 2026-09-30