Anthropic Model Reasoning Traces Accessible for Months, Raising Privacy Concerns

jessi_cata · x · 2026-08-12

A tweet highlights that reasoning traces from Anthropic models have been accessible for months, raising privacy concerns. Users worry that even if outputs don't contain private data, the reasoning process may access it, and sharing chats could expose these traces.

Related event: Researchers Extract Hidden Reasoning Traces from Proprietary LLMs(35 posts)→

Original post →

More from Safety

Safety channel →