Additional Insights on Inkling Efficiency and Scores
ArtificialAnlys · x · 2026-07-16
This reply shares further key benchmark data for Inkling:
- GDPval-AA v2 Elo is 1238, beating Kimi K.6 and DeepSeek v4 Flash max
- Token efficiency: Averages around 25K output tokens per Intelligence Index task, fewer than GLM-5.2, Kimi K2.6, and DeepSeek v4 Pro
This remains a breakdown of Inkling's overall performance.
More from Models
- Counterfactual: without reasoning models, AI today might just be reaching o3-level — Jsevillamol · 2026-09-03
- Anthropic's Fable 5.1 hits 90% on ARC-AGI-2 at 32% lower cost per task than Fable 5 — rohanpaul_ai · 2026-09-03
- Team shares 4 real LLM uses: contract negotiation, agent clarification, grading, math — xuanalogue · 2026-09-03
- Team claims h3 max is the undisputed #1 frontier video model across benchmarks — isidentical · 2026-09-03
- Meta's SAM 3, with image and video segmentation, tops Hugging Face trending — facebook · 2026-09-03
- Anthropic internal 'retirement home' for old Claude models sparks confabulation concerns — repligate · 2026-09-03