MiniCPM5-2B lands on Hugging Face as author claims Llama architecture beats Qwen3.5
solyarisoftware · x · 2026-09-12
openbmb has released MiniCPM5-2B on Hugging Face. Benjamin Marie's takeaway: based on this model, the Llama architecture can outperform Qwen3.5, suggesting architecture choice still matters for small models. The model card adopts a Qwen-style chat template with tool-calling support.
More from Models
- Codex quality fixes shipped, reset rolling out to ChatGPT Work users — CtrlAltDwayne · 2026-09-12
- Codex usage reset reportedly rolling out Sep 12 at 07:00 UTC, execution unconfirmed — DevDminGod · 2026-09-12
- Qwen3.8-27B goes live on Cerebras with fast inference, scoring 34 on AAII — Alibaba_Qwen · 2026-09-12
- antirez uploads DeepSeek V4.1 Flash GGUF quantizations to Hugging Face — Queasy_Asparagus69 · 2026-09-12
- Qwen3.8-Max scores 19 on OpenRouter vs 26 on Alibaba — effort level bug suspected — PawelHuryn · 2026-09-12
- Pipecat v1.9 adds Meta's Muse Voice Transcribe, the lowest semantic-WER STT model tested — solyarisoftware · 2026-09-12