GPT-6-luna Unlocks More Reasoning Tokens via API: ~18k Tokens Scores ~80.5% on Terminal-Bench

LysandreJik · x · 2026-10-07

Testing by evalstate shows GPT-6-luna delivers more reasoning tokens per effort level when accessed with an API key — worth knowing if you call it through different routes/providers. Using Codex CLI 0.160.0 with terminal-bench 2.1 (excluding 10x QEMU tasks): for tb-21 tasks, 10.5k output tokens is the point of diminishing returns; API max reaches 18,000 tokens at a score of 80.5%, just below Luna 5.6 in the same testing.

Original post →

More from coding & agent

coding & agent channel →