Stealth Model 'Ox Alpha' Matches GPT-5.6 in Context Arena Tests
scaling01 · x · 2026-08-24
A new stealth model named "Ox Alpha" has been added to the Context Arena, running via OpenRouter. It achieves a 75.3% AUC@128k score on high thinking settings, matching Gemini 3.1 Pro and GPT-5.6 Luna while beating Claude Sonnet 5. Interestingly, enabling max thinking reduces accuracy beyond 32k context length despite using 34% more tokens. The model maintains a 39.0% score at 512k context length.
More from Models
- Mystery OxAlpha Beats Claude; Alibaba Raises $10B for AI — 创业邦 · 2026-08-24
- OpenAI and Google cut LLM prices; mystery OxAlpha model beats Claude on DeepSWE — 创业邦 · 2026-08-24
- AI News Digest: DeepSeek Weekend Discounts, GPT-5.6 Sol Price Cut, Alibaba's $10B AI Raise — APPSO · 2026-08-24
- Fable 5 Makes Up Only 6% of Anthropic Sales — rohanpaul_ai · 2026-08-24
- Opus 5 writes like Sorkin? How to fix wordy style — thk_ · 2026-08-24
- 2026 Chinese Model Landscape: DeepSeek V4, Kimi K3, and More Listed — TheTuringPost · 2026-08-24