Using Pangram as a writing yardstick may degrade model quality

max_spero_ · x · 2026-08-19

Max Spero argues against using the AI text detection tool Pangram as a benchmark for LLM writing quality, fearing it will harm output. He also states that training undetectable LLMs is detrimental to society.

Related event: Critics warn AI detectors are a bad proxy for writing quality(2 posts)→

Original post →

More from Safety

Safety channel →