Ling-3.0-flash-Fin full benchmark results: ~67k output tokens per task
ArtificialAnlys · x · 2026-09-16
Artificial Analysis published full results for Ant Group's Ling-3.0-flash-Fin across all 10 evaluations in the Intelligence Index v4.3. A notable finding: the model is verbose, averaging 67k output tokens per task — 34% more than Ling-3.0-flash-VL (50k) and over 3x MiniMax-M2.7 (21k).
Related event: Ant's Open-Source Ling-3.0-flash-Fin Benchmarked by Artificial Analysis(3 posts)→
More from Models
- Users report ChatGPT sessions getting muddled, answering questions from other chats — koltregaskes · 2026-09-16
- OpenAI reportedly prepping Codex Replay to run and compare historical task threads in parallel — testingcatalog · 2026-09-16
- Mystery stealth model Union Alpha hits OpenRouter: free, 256K context, agentic focus — gaganghotra_ · 2026-09-16
- Developer builds demo hours after getting Jev access, drawing researcher banter — suchenzang · 2026-09-16
- TabPFN-3.5 tops Kaggle's Otto competition out of the box, but experts call the benchmark flawed — RichmanRonald · 2026-09-16
- Bindu Reddy teases Opus 5.2 in testing, Grok 4.8 weeks away, OpenAI's Astra+ in testing — bindureddy · 2026-09-16