Dev Fine-Tunes Qwen3.8-27B on 125K Real Conversations to Kill the AI Assistant Vibe

kvyb · reddit · 2026-09-12

An indie developer released Qwen3.8-27B-Humanlike-Chat, a LoRA fine-tune (rank-256, checkpoint 863) trained on 125,217 obfuscated human-to-human messages across 1,396 conversations, aimed at removing the 'AI assistant' tone — replies come out shorter, less polished, more human, without a system prompt. The tradeoff: an earlier checkpoint scored 5 points lower on IFEval, and coding wasn't tested. Ships with merged GGUFs, an HF Space demo, side-by-side comparisons, and a free rate-limited OpenAI-compatible API.

Original post →

More from Models

Models channel →