ARC Prize finds Qwen3.8-27B's chat template injects different instructions per reasoning effort

GregKamradt · x · 2026-10-02

Testing Qwen3.8-27B via Baseten, the ARC Prize team discovered the model's own chat template injects different instructions per reasoning effort: low tells it to keep thinking brief, xhigh instructs careful thinking and checking assumptions, while medium adds neither — even though thinking stays enabled. This may explain medium's lower scores: the settings change how the model is instructed to approach problems, not simply its thinking budget. GregKamradt called it unexpected and credited Baseten for verifying the finding; the template is public on Hugging Face.

Original post →

More from Models

Models channel →