llama.cpp ignores reasoning_effort for Qwen3.8, breaking reasoning control
fbms2 · reddit · 2026-08-15
llama.cpp's llama-server only handles reasoningeffort='none', ignoring other values, so Qwen3.8-27B's chat template can't adjust reasoning. The user provides root cause and a fix, linking related issues.
More from coding & agent
- Designing Payment Authorization for AI Agents: Balancing Security and Autonomy — NoCalendar831 · 2026-08-15
- Ex-Meta Scientist: Agents should use the web like humans via pixels and clicks — DhruvBatra_ · 2026-08-15
- OpenAI Codex Error: 'gpt-5.6-sol' Model Possibly Deprecated — AKsnipebuster47 · 2026-08-15
- reBot Arm Control Stack Integrates Agentic AI, VLM, and LLM for Robotics — kamathsblog · 2026-08-15
- Dev's Cloud Agent experience: escaping pesky permission prompts — davidcrawshaw · 2026-08-15
- Tested 3 models to spec a local AI-brain install: one cited real files, one got macOS compat backwards — schwentker · 2026-08-15