Thomson-1.0-Small: A RAG-Optimized Fine-tune of Qwen
uber-linny · reddit · 2026-08-28
The author recommends Thomson-1.0-Small, a fine-tune based on Qwen, as a better alternative for RAG and document review than the deprecated 9B/35B MOE models. It runs at 25 T/s on a 9070xt with good quality.
More from Models
- GLM-5.2 Introduces Monitors to Combat Reward Hacking in RL — burny_tech · 2026-08-28
- Anthropic Luna Max test shows generous limits, high speed — timpera · 2026-08-28
- User calls Grokbot 'terrible', cites missing tasks — krishnan · 2026-08-28
- Prime Intellect Evaluates Autonomous AI Research Capabilities Across 18 Frontier Models — mariofilhoml · 2026-08-28
- Optimize Models to Think Less, Not Just Generate More Reasoning Tokens — abacaj · 2026-08-28
- ChatGPT keeps appending mysterious code to user chats — TheMoonMidas · 2026-08-28