gpt-6-luna lands #25 on decision index at 2.3x the cost, researcher says
antoine_chaffin · x · 2026-10-07
antoinechaffin argues that paying so much for so little is absurd: a decision model (non-generative classification) should not be a wrapper on top of generative models. Per multimodalart, gpt-6-luna ranks #25 on the jev decision index v0.3 at 2.3x the cost of jev. His own model, which beats jev, is open — runnable locally or via a cheap API.
More from Models
- Temp 0 doesn't guarantee deterministic LLM output — batch shape and kernels shift logits — JFPuget · 2026-10-07
- OpenAI doesn't need more resets, needs better communication, argues paying user — AirportEither2456 · 2026-10-07
- GPT-6-luna Unlocks More Reasoning Tokens via API: ~18k Tokens Scores ~80.5% on Terminal-Bench — LysandreJik · 2026-10-07
- Bug Hunt Benchmark retest: GPT-6.1 Sol recovers, Muse still cheapest strong agent — PawelHuryn · 2026-10-07
- Subscription multipliers recalculated: SuperGrok gives 80x, Muse Power+Contributor 137x usage — PawelHuryn · 2026-10-07
- r/mathematics Reacts to OpenAI's New Math Proofs, Takes Are Split — petburiraja · 2026-10-07