Anthropic resumes billing for safety-blocked requests; 99.7% of users unaffected
ClaudeDevs · x · 2026-09-25
Anthropic says it will resume charging for requests its safeguards block before Claude responds, limited to low-false-positive categories (biology, distillation attacks, frontier LLM development), citing coordinated attacks. It reports 99.7% of users won't hit these billable blocks, classifiers are tuned to <0.1% false positives, and misfires can be reported via /feedback in Claude Code.
More from Models
- Predicting a new GPT-4 / Sonnet 3.5 / Opus 4.5 moment from upcoming GPT and Claude releases — kieranklaassen · 2026-09-25
- Anthropic engineer: crank Claude's effort setting for risky code, dial it down for routine tasks — every · 2026-09-25
- Proteus, accepted at NeurIPS 2026, unlocks memory blocks to boost long-context by +8.4 NIAH — behrouz_ali · 2026-09-25
- LangChain launches LangSmith Fine-Tuning with smithtune CLI to train models from agent traces — LangChain · 2026-09-25
- xAI's "xhigh latest" on Pro tier panned as merely "Qwen-tier" — teortaxesTex · 2026-09-25
- AIRA₂ Research Agents Hit 81.5% on MLE-bench-30, Beating Prior SoTA of 72.7% — mariofilhoml · 2026-09-25