Anthropic's ban system criticized: no human reviews make it risky for businesses

jdjohnson · x · 2026-10-09

Responding to the argument that Anthropic's politeness policy is good for training data quality, jdjohnson counters that harmful language should simply be filtered from training data, and that Anthropic's automated banning system is broken with no human review to fix false positives. As the ban net widens, more innocent users get caught, making Anthropic a risky choice for businesses that lack the leverage to obtain human review.

Original post →

More from Companies & People

Companies & People channel →