Qwen Flash Occasionally 'Thinks Into the Air' After Tool Calls, User Reports
pilibitti · reddit · 2026-09-24
A user reports that Qwen Flash (both Unsloth and AtomicChat quants) occasionally produces odd, self-referential thinking blocks after tool calls — reasoning about system instructions as if no user task exists — before continuing normally. Cause unknown; possibly related to message-history handling.
More from Models
- A private eval with a 0% completion rate for 3 years: no AI model can identify this flag — generativist · 2026-09-24
- Contrastive-LM org ships CLM-v0.1-8B model and Nemotron pretraining dataset on HF — _akhaliq · 2026-09-24
- Former OpenAI VP Brundage: ChatGPT web keeps resetting voice from Astra to Sol — Miles_Brundage · 2026-09-24
- Models now generate animated videos from scratch via HTML and Blender faster than predicted — xuanalogue · 2026-09-24
- Opus 5.5's safety classifier refuses to visualize another model's rollouts — eliebakouch · 2026-09-24
- Ex-OpenAI team launches Jev, a non-LLM model that outputs typed values with confidence scores, raises $40M — cephaloform · 2026-09-24