Sapience System Reads Only ~15K Tokens Yet Beats GPT-6 Astra on MRCR, 97.7 vs 96.3
gordic_aleksa · x · 2026-09-15
A surprising long-context result: Sapience, a custom system that reads only 15K tokens and delegates answering to open model DeepSeek V4 Pro, scores 97.7 on OpenAI's MRCR benchmark — beating GPT-6 Astra's 96.3, which must ingest the full 500K–1M token context.
The author notes similar gains on RULER and other long-context benchmarks, suggesting it's not overfit to a single benchmark — hinting that smart retrieval may beat brute-force full-context reading.
Related event: Sapience tops GPT-6 Astra on MRCR reading just 15K tokens(2 posts)→
More from Models
- Physical Intelligence unveils OM-1, a robot foundation model trained purely on human data — zipengfu · 2026-09-15
- AI-text detector Pangram is powerful but badly calibrated, dev argues — maksym_andr · 2026-09-15
- Claude Code's 50% promo ended, users get 17% less usage — msg · 2026-09-15
- Running Qwen3.8 Flash Next on 128GB RAM + one 5080: 136pp/19tg at Q5_K_XL — whatyathinkk · 2026-09-15
- DeepSeek-V4.1-Flash Hits #3 Open Model on Agent Arena at $0.07 per Task — arena · 2026-09-15
- Researcher: RL Bar Has Been Raised Due to Reward Hacking, More to Come — tszzl · 2026-09-15