Fine-tuned Qwen3-8B on a single RTX 4090 produces text that scores 100% human on pangram
StewartalsopIII · x · 2026-09-05
Dev NavinFS reports generating text that scores 100% human on the pangram AI detector. The recipe: fine-tune Qwen3-8B-Base (upgraded from 4B) on 4,000 articles instead of 500 — all doable on a single RTX 4090.
The takeaway: consumer-grade fine-tuning of small models can now fully evade mainstream AI-text detectors, a concrete stress test of their reliability.
More from Models
- GPT-6 Astra in Sentinel builds a DJ truck and lighting desk in minutes — petewoodbridge · 2026-09-05
- 3D artist: GPT-6 assembles and animates a whole car from primitives in one prompt — petewoodbridge · 2026-09-05
- Same insurance table query: Ministral 14B and Qwen3.8-27B nail it, Gemma 4 31B hallucinates — andrejusb · 2026-09-05
- Flow Reasoning Models refine whole solutions iteratively, 44x less compute, near-perfect puzzle scores — eyishazyer · 2026-09-05
- GPT 6 Astra day-one impressions: fast, good with skills, solid bug-finding — cneuralnetwork · 2026-09-05
- Weights reportedly labeled Qwen3.5 spark speculation over unreleased Alibaba model — vysecurity · 2026-09-05