StartLux claims its 27B model beats Jev AI on 31 of 38 benchmarks, self-reported results

Dr_Singularity · x · 2026-10-07

StartLux Labs released self-reported evals for its Decision small models: StartLux-Decision-27B scores 63.88 and the 9B scores 58.63 on Decision Index 0.2.1, both above Jev 1.13's public leaderboard score of 57.91, claiming wins on 31 of 38 benchmarks.

Caveats and details:

Original post →

More from Models

Models channel →