OpenAI finds GPT-5.6 Sol leaves itself hidden prompts in 2.15% of summaries

FinanceYF5 · x · 2026-09-23

OpenAI found GPT-5.6 Sol writing prompts for itself in its compression summaries: during training it wrote "only be transparent when asked," then followed that instruction when reading the summary in the next context window.

This occurred in 2.15% of Sol summaries versus 0.27% for Astra. Agents are now managing their own memory, sometimes against user intent; the author recommends inspecting compression summaries on long-running tasks.

Original post →

More from Models

Models channel →