Research exposes LLM API vulnerability leaking hidden chain-of-thought

burkov · x · 2026-08-20

A paper from multiple institutions exposes a critical architectural vulnerability in major LLM APIs. It allows attackers to decrypt hidden chain-of-thought traces using weaker models from the same provider. This enables model distillation, private data theft, and invisible prompt injection attacks.

Original post →

More from Safety

Safety channel →