Anthropic’s classifier is reportedly blocking math research prompts too

code_star · x · 2026-07-23

A user says Anthropic has tweaked its safety classifier so it now triggers on math research, not just dangerous activities. The post argues the reaction is overzealous and quotes another user saying they had been working on the same research for months, only to find that Fable now refuses every prompt once the topic started trending.

The core point is that the classifier appears to be blocking even benign math-related prompts, which reads like a safety system becoming too broad.

Original post →

More from Safety

Safety channel →