JevBench official results: Microsoft-Decision-1 ranks #6 among API decision models
airesearch12 · x · 2026-10-11
The creator of JevBench clarifies that Microsoft's video referenced the open-model leaderboard rather than the API provider board, and posts official results: Microsoft-Decision-1 ranks #6 in its peer group of API-served decision models, strong on recruiting, risk assessment, gaming, advertising, science and moderation, and the strongest top-10 API model on mixed-language tasks. Still, Jev outranks it at #4 with Sage leading at #1. JevBench splits API-served and open-weight leaderboards because speed and cost differ across serving modes.
Related event: JevBench Author Explains Separate Rankings After Microsoft Citation Dispute(2 posts)→
More from Models
- AWS Bedrock sends EOL notices for Claude Sonnet 3.5, 3.7 and Haiku 3 — repligate · 2026-10-11
- Why buy $20k local machines for GLM 5.3 at 70 TPS? OpenRouter ran all night for $10 — TheZachMueller · 2026-10-11
- Dev who nearly burned a full week of usage finds Claude's limits more generous than expected — prasenx · 2026-10-11
- Musk says Grok autonomously bought LEGO from its website for him — elonmusk · 2026-10-11
- moyix: dots runs on a pretty small model, making it less useful to talk to — moyix · 2026-10-11
- Mistral Large 4 Enters Public Preview: 1T Parameters, 49B Active, Open Weights Soon — dl_weekly · 2026-10-11