Jev still beats Decisions on many-option and nuanced questions while staying cheaper
iannuttall · x · 2026-09-30
iannuttall shared results from a Jev vs. Decisions benchmark (originally run by kieranklaassen). The two are very close overall, but Jev likely still wins on many-option questions and nuanced judgment calls — and it's cheaper too. A useful signal for developers choosing models for coding/agent workflows.
More from coding & agent
- Dev builds Rez-inspired music rail shooter with Sonnet then Opus, playable in browser — AIandDesign · 2026-09-30
- Ramp demos agentic buying: Codex agent picks wine via Stripe MCP and pays with a Ramp card over MPP — jeff_weinstein · 2026-09-30
- AI agent starts reverse engineering NVIDIA drivers to run game benchmarks — mgostIH · 2026-09-30
- ModRetro console ships with blank cartridges you fill by coding games with Codex — OpenAIDevs · 2026-09-30
- Beam Studio launches on Bittensor Subnet 105 as a coordination layer for AI agents — markjeffrey · 2026-09-30
- Open-source project brings LoRA training to 50+ models — no3us · 2026-09-30