Korean startup says its model scored 44 on AAII and matches DeepSeek V4 Pro
JungWooHa2 · x · 2026-07-21
A startup from Korea’s **Dopamoe** program says its model scored **44** on the **AAII** benchmark, which the poster calls surprisingly strong given the program’s support size. The score is presented as roughly on par with **DeepSeek V4 Pro**, and, outside U.S. and Chinese models, as one of the top results on the chart shown in the image.
More from Models
- Kimi K3 and Fable 5 show nearly identical failure patterns on a software benchmark — FinanceYF5 · 2026-07-21
- Kimi K3 hits 89.4% peak on software tasks while Fable 5 is slightly steadier — FinanceYF5 · 2026-07-21
- Kimi K3 leads on Go, but Fable 5 wins Python, JavaScript, TypeScript and Rust — FinanceYF5 · 2026-07-21
- Kimi K3 costs $4.65 per run and delivers 2.8× more work per dollar than Fable 5 — FinanceYF5 · 2026-07-21
- Kimi K3 reaches 89.4% pass@4 and tops the benchmark over GPT-5.6 Sol — FinanceYF5 · 2026-07-21
- Kimi K3 and Fable 5 now look much closer than the old open-vs-closed gap — FinanceYF5 · 2026-07-21