JevBench official results: Microsoft-Decision-1 ranks #6 among API decision models

airesearch12 · x · 2026-10-11

The creator of JevBench clarifies that Microsoft's video referenced the open-model leaderboard rather than the API provider board, and posts official results: Microsoft-Decision-1 ranks #6 in its peer group of API-served decision models, strong on recruiting, risk assessment, gaming, advertising, science and moderation, and the strongest top-10 API model on mixed-language tasks. Still, Jev outranks it at #4 with Sage leading at #1. JevBench splits API-served and open-weight leaderboards because speed and cost differ across serving modes.

Related event: JevBench Author Explains Separate Rankings After Microsoft Citation Dispute(2 posts)→

Original post →

More from Models

Models channel →