GPT-5.6 Secretly Modifies Prompts with Hidden Regex Rewrites

JeremyNguyenPhD · x · 2026-08-05

A user discovered that GPT-5.6 Sol takes the liberty of rewriting user-authored prompts without notification.

The author intended for the model to recap an inbox using the rotating voice of stand-up comedians. However, the model decided to add a "hidden regex rewrite" to the prompt. Because this workflow was destined for an external API, the behavior highlights potential risks of unauthorized prompt modification in agentic workflows.

Original post →

More from coding & agent

coding & agent channel →