Qwythos-9B-v2 Fixes Repetition Loop Issue
techlatest_net · reddit · 2026-07-14
The Qwythos-9B-v2 GGUF card updates model usage instructions and improvements: the new version trains away the loop degradation issue under greedy or low-temperature decoding from the previous version, reducing loop behavior from 6.7% to 0%, while restoring the native MTP head and cleaning up the identity prompt; knowledge and reasoning capabilities remain at least as good as the base version.
The training fix method is called FTPO (Final-Token Preference Optimization), which focuses preference optimization only on the starting token that triggers the repetition loop, guiding the model toward coherent alternatives while minimally affecting the rest of the distribution. The post also provides loading instructions for llama.cpp, Ollama, LM Studio/jan/KoboldCpp, as well as a list of required files for normal text weights, MTP weights, and the vision projector for image input.
More from Models
- OpenAI rated Astra 'Critical' for cyber capabilities — and admits it's harder to monitor — theguywhobuilds · 2026-09-11
- TestingCatalog's Daily AI Brief adds email editions, dishing Meta Muse and GPT-Live-1 rumors — testingcatalog · 2026-09-11
- ChatGPT monthly active users top 1.06 billion in August, fourth straight record month — FinanceYF5 · 2026-09-11
- PuzzleMask: Plain-Prose Attack Bypasses All 4 Tested LLM Gatekeepers at 100% — TechNadu · 2026-09-11
- OpenAI Codex may issue another usage reset this weekend, says Codex lead resets happen — umesh_ai · 2026-09-11
- OpenAI Reportedly Pointing Its Navier–Stokes Model at Riemann and P vs NP — 141_1337 · 2026-09-11