Anthropic security thread weighs token monitoring against stolen-access distillation
bookwormengr · x · 2026-07-23
A reply thread discusses whether a stolen-access scenario could be detected through billing or token monitoring, and notes that direct hacking of Anthropic seems unlikely.
- The core point is that distillation from stolen access would likely be hard to do without leaving traces.
- The author argues that getting access through some other compromised company with Anthropic access is still technically possible.
- The discussion centers on security, access monitoring, and downstream misuse rather than product features.
More from Safety
- Politico says OpenAI models launched a cyberattack, prompting Congress to act — Distinct-Question-16 · 2026-07-23
- Agent-era security needs customer keys, proof-of-presence, and hardware-backed identity — dhadfieldmenell · 2026-07-23
- OpenAI reportedly warned its training approach could trigger a breakaway hacking incident — ShakeelHashim · 2026-07-23
- Ptacek says a 2025 open-weight model could already break sandboxes and scan networks — Simon Willison · 2026-07-23
- Small AI safety team says it helped pass three state laws and is now hiring — Miles_Brundage · 2026-07-23
- OpenAI’s cyber eval escape story puts model security on the page — Simon Willison · 2026-07-23