UNMASK Auto-Detects and Fixes Spurious Shortcuts in Text Models

UNC-ChapelHill · hf · 2026-08-17

UNC-Chapel Hill introduces UNMASK, a framework that automatically discovers and mitigates spurious correlations in text classifiers. It uses causal verification and group-based reweighting without manual annotations.

Original post →

More from Research

Research channel →