HarmProfile Benchmark: Harmfulness and Diversity Rise with Model Capability

Zhouyuan Ma · hf · 2026-08-19

HarmProfile is a benchmark dataset characterizing frontier LLM safety failures through content analysis. It reveals that harmfulness and diversity increase with model capability.

Original post →

More from Safety

Safety channel →