Dev Fine-Tunes Qwen3.8-27B on 125K Real Conversations to Kill the AI Assistant Vibe
kvyb · reddit · 2026-09-12
An indie developer released Qwen3.8-27B-Humanlike-Chat, a LoRA fine-tune (rank-256, checkpoint 863) trained on 125,217 obfuscated human-to-human messages across 1,396 conversations, aimed at removing the 'AI assistant' tone — replies come out shorter, less polished, more human, without a system prompt. The tradeoff: an earlier checkpoint scored 5 points lower on IFEval, and coding wasn't tested. Ships with merged GGUFs, an HF Space demo, side-by-side comparisons, and a free rate-limited OpenAI-compatible API.
More from Models
- Kording Lab: AI credits inventors but almost never credits the original source of an idea — KordingLab · 2026-09-12
- DeepSeek Flash 4.1 Draws Early Praise From Developer Who Tried It — _arohan_ · 2026-09-12
- The real worry in Anthropic's report: no way to stop distillation by rivals — kimmonismus · 2026-09-12
- Meta's Confidential Computing Promise Doesn't Protect Data if Inference Isn't Attested — signulll · 2026-09-12
- Among a Flood of New AI Labs, the Humans Team Ships Its First Release — niloofar_mire · 2026-09-12
- Agent traces reveal hour-long costly 'super resolution' rabbit hole in continual learning study — yuxiangw_cs · 2026-09-12