Ling 3.1 Flash supplemental scores: Terminal-Bench 0%→33%, hallucination 38%
ArtificialAnlys · x · 2026-10-06
Additional Artificial Analysis data on Ling 3.1 Flash: Terminal-Bench v4.0 rises from 0% to 33%, AutomationBench-AA from 3% to 62%; AA-Omniscience scores +2 with 29% accuracy and a 38% hallucination rate.
Related event: Ant Group's Ling 3.1 Flash Doubles Intelligence Index to 41(4 posts)→
More from Models
- GPT-6.1 Sol debuts at #5 on PostTrainBench, behind GPT-6 Astra and Opus 5.5 — maksym_andr · 2026-10-06
- ChatGPT can't decide which language to title its chats in — cool101wool · 2026-10-06
- GLM 5.3 Flash runs locally on dual V100s: 93GB MoE weights at ~20 tokens/s — lxfater · 2026-10-06
- Qwen3-Next-80B on a 3090 with 16GB RAM: ~3x decode speedup via MoE expert substitution — Zestyclose_Reality15 · 2026-10-06
- Anthropic investigating elevated errors on Claude Opus 5.5 requests — ClaudeAI-mod-bot · 2026-10-06
- ARC-1: a 1.7B decision model answering in ~20ms on a 4060 Ti, free and local — KMatysek · 2026-10-06