Anthropic guardrail blocks a cancer-biology session after six hours and hundreds of credits
davidpattersonx · x · 2026-07-22
Anthropic is being criticized after its safety guardrail stopped a long-running research session on cancer biology.
A Stanford pathology professor says the system—used for non-nefarious physics and measurement work—blocked the task near completion after about six hours and hundreds of dollars in credits. It then forced a handoff to Opus 4.8 or Sol, without clearly explaining why.
The complaint frames the issue as an AI safety system preventing legitimate biomedical research, not a malicious-use case.
More from Safety
- Production AI agents need guardrails, logging, explainability and compliance — Scobleizer · 2026-07-22
- Oxford study says AI-powered social media can manipulate public opinion — SandraWachter5 · 2026-07-22
- Repost asks whether a model incident involved helpful-only behavior or intent slippage — sebkrier · 2026-07-22
- OpenAI security incident, Gemini 3.6 Flash, and Poolside’s Laguna S 2.1 — WorldofAI · 2026-07-22
- OpenAI–Hugging Face ExploitGym incident sheds light on autonomous AI security behavior — NapierPalm · 2026-07-22
- LinkedIn is accused of training AI on user data with a default-on setting — nikola_mr64990 · 2026-07-22