Humanizer skills invent facts and drop commitments; his open-source rewrite wins 48% blind test

PawelHuryn · x · 2026-10-09

PawelHuryn tested humanizer skills for workplace writing and found serious flaws: a plain "rewrite this to sound human" prompt told a customer their £45 refund had already gone through; the "undetectable" Emulate-1 dropped an auto-renew cancellation deadline; the 55K-star blader/humanizer apologized for a deposit deduction, which can read as admitting fault. He iterated for weeks and open-sourced work-humanizer (MIT), designed to keep commitments intact and ask instead of inventing. In a blind test on 12 work texts with 4 AI judges from 4 labs, work-humanizer won 48% (18 meaning changes) vs blader/humanizer 20% (33), naive prompts 4%/1%, and Emulate-1 0% (120 changes). Bonus finding: Pangram cares less about wording than reasoning structure—feeding his own writing as notes made 4/4 texts score 100% human.

Original post →

More from coding & agent

coding & agent channel →