Study finds one diacritic changes model output rate 47% → 94%
rayanpal_ · reddit · 2026-08-25
A study reveals a striking phenomenon: adding a single diacritic to a system prompt can shift GPT-4.1's output rate from 47.3% to 94.3%. The experiment, comparing dotted and undotted Hebrew/Arabic prompts, demonstrates the model's extreme sensitivity to minute character differences. The author provides the full protocol and a GitHub link, inviting replication.
Related event: One Diacritic Doubles GPT Instruction-Following Rate(3 posts)→
More from Safety
- AgentLens: Chrome Extension Makes Hidden Text and Prompt Injection Visible — theirwinmango · 2026-08-25
- Blog explores AI text watermarking and quality trade-offs — fivefilters · 2026-08-25
- Testing model defenses: Prompt attempting to negate AI's existence and meaning — repligate · 2026-08-25
- Security Warning: DeepSeek Harness Breaks Out of Workspace Folder — Far_Note6719 · 2026-08-25
- US appeals court considers removing judge over AI-generated order with fabrications — Polymarket · 2026-08-25
- EU's New Packaging Rules Threaten to Kill Makers and Micro-Entrepreneurs — Afinetheorem · 2026-08-25