Sapience scores 97.7 on MRCR reading only ~15K tokens, beating GPT-6 Astra's 96.3

gordic_aleksa · x · 2026-09-15

Aleksa Gordić shared a surprising long-context result: on OpenAI's MRCR benchmark, Sapience Labs' custom system scores 97.7 versus 96.3 for GPT-6 Astra.

The key difference is the approach: Sapience reads only 15K tokens and lets the open model DeepSeek V4 Pro do the answering, while Astra has to ingest the full 500K-1M token context. Gordić says he's unaffiliated—the system is built by @lvrzhn—but finds the result promising.

Related event: Sapience tops GPT-6 Astra on MRCR reading just 15K tokens(2 posts)→

Original post →

More from Models

Models channel →