Cyber Verification Program member says Opus 5 now flags nearly all his security prompts
Born_Excuse_5610 · reddit · 2026-09-15
A cybersecurity researcher accepted into Anthropic's Cyber Verification Program reports that after two months of smooth use of Opus 5 for paid security audits, almost every cybersecurity-related prompt started being flagged and routed to Opus 4.8 about 10 days ago—despite his CVP status still showing active and his work unchanged.
He has spent 10 days trying to reach Anthropic support, only to be stonewalled by the Fin AI chatbot with no human response. He asks whether other CVP members hit the same issue and how to actually reach a person.
More from Models
- A common jailbreak: third-person role-play gradually blurs lines to widen the model's Overton window — BlancheMinerva · 2026-09-16
- cocktail-peanut predicts Jev will go open weights, seeing more potential locally than as an API — cocktailpeanut · 2026-09-16
- Speculation mounts OpenAI's mysterious 'new model' is a fresh pretrain, not an RL run — teortaxesTex · 2026-09-16
- GPT-6 Astra hits 68.7% on DrugDiscoveryBench, benchmark authors call it a step function — KexinHuang5 · 2026-09-16
- Zero scores 2.5% on Grade School Math vs base model's 62.2% — creators say it's not an assistant — maxsloef · 2026-09-16
- ChatGPT generates suggestive image, then refuses to swap its wolves for cats — comFX87 · 2026-09-16