NSA advisory urges silent downgrades for suspected distillers, clashing with Anthropic's June transparency promise

MysteriousAvocado580 · reddit · 2026-10-05

A detailed Reddit post examines the tension between Anthropic's June apology and the Sept 8 NSA/CISA/FBI advisory AA26-251A:

The core problem is detection: listed indicators like 24/7 usage without idle periods, max-usage new subscriptions, and cache-optimization-tuned traffic are a fair description of any well-tuned scheduled agent fleet. The author, running a small multi-agent setup, was hit twice by reasoningextraction refusals on legitimate work (a text-leak bug discussion) — only caught because the refusal carried a label. A silent downgrade would have looked like "a bad day."

Contrasts: Anthropic's transparency page shows 11.4M banned accounts with appeal paths; OpenAI's Sept 30 shutdown of a Moonshot-linked reasoning-replay campaign used visible bans; Fable 5.1's anti-distillation tightening is visible too. But the author can find no statement since Sept 8 on whether "every time it happens" still holds.

Original post →

More from Models

Models channel →