Mimir: 1.7B model claims to beat Qwen and Gemma
ZookeepergameCool173 · reddit · 2026-08-17
Introduction of a new small model, DFM-Mimir (claimed 1B, approx. 1.7B parameters). It claims to outperform Qwen 2.5 0.8B/2B and Gemma 2 2B on various benchmarks. The model supports English and Danish, with decent math and coding performance. It is built on Sapient's hrm-text architecture using layer-recurrence techniques.
More from Models
- Local Object Detection with Qwen3.8-27B: Cross-Validating with RF-DETR — MaziyarPanahi · 2026-08-17
- DFM Mimir v1: Open 1B Model Achieves SOTA Danish Performance — SDU-Denmark · 2026-08-17
- Ling-3.0-flash Runtime Path: Running on One DGX Spark — Kanu-animallover · 2026-08-17
- System prompts beat user prompts: taming verbose Claude Opus 5 — IndyDevDan · 2026-08-17
- Intern-S2-Mobius: Decoupled Knowledge and Reasoning Model — pmttyji · 2026-08-17
- Tencent's EVIE Model Tops ViDoRe Benchmarks, Cuts Vector Storage Costs by 32x — jacek2023 · 2026-08-17