Nanbeige 4.2 3B is producing garbled reasoning output in llama.cpp
LaurentPayot · reddit · 2026-07-28
Nanbeige 4.2 3B is spitting garbled reasoning output in llama.cpp
The user reports that Nanbeige 4.2 3B, running on the latest llama.cpp, produces broken reasoning traces and nonsensical output even for a simple “hi” prompt. Two different agent setups behave badly: one emits repeated </think> fragments and random text, while another seems to loop forever.
The post includes the full launch command with --reasoning on, --reasoning-preserve, a 65,536 context window, and quantized KV cache settings. It is essentially a model/runtime bug report asking whether others are seeing the same issue.
More from Models
- SingularityAPI bundles DeepSeek and Kimi models behind one OpenAI-compatible endpoint with free beta credits — Individual_Team_2344 · 2026-07-28
- Rising infra costs make custom chips tempting, but Nvidia is still a hard fight — FinanceYF5 · 2026-07-28
- Reddit user says Claude Pro now hits its session cap after 1–2 prompts — Automatic-Ad-6936 · 2026-07-28
- Claude screenshot turns Anthropic’s naming confusion into a running joke — heypearlai · 2026-07-28
- Hundreds of Claude chats were indexed publicly after users shared them — connoraxiotes · 2026-07-28
- Unsloth’s Kimi-K3-GGUF starts trending on Hugging Face as an image-text model package — unsloth · 2026-07-28