Mythos Preview Leads in Long Context
scaling01 · x · 2026-07-17
This post relays a model comparison result: **Mythos Preview** outperforms **GPT-5.6-Sol** and **Mythos 5** in the first approximately **30 million token** long-context phase. The author adds that Mythos Preview's best attempt speed is also slightly faster than that of GPT-5.6-Sol. A cited comment provides another comparison: in UK AISI's long-range cyber attack and defense tests, **open-models** lag behind frontier models by about **7 months**. Notably, **GLM-5.2** is closing in on **Opus 4.5** on the `The Last Ones` cyber range, and matches **Opus 4.6** on shorter, narrower tasks.
More from Models
- Claude 20x users report sharply tighter limits and faster quota burn — MarcJSchmidt · 2026-07-21
- Cola launches July, the latest model jokingly billed as “second only to Fable” — oran_ge · 2026-07-21
- Kimi K3 looks stronger and about 5× cheaper on a frontend dashboard task — OwariDa · 2026-07-21
- Last Week in AI recap: Anthropic’s $65B round, IPO filing, and Microsoft’s MAI push — Last Week in AI · 2026-07-21
- A user says Claude 4.6 felt worse yesterday and asks whether model quality can drift over time — Rahios · 2026-07-21
- Kimi K3 hits 89.4% peak on software tasks while Fable 5 is slightly steadier — FinanceYF5 · 2026-07-21