PhD thesis examines where automatic NLG evaluation fails by domain

mdredze · x · 2026-07-21

A PhD student has defended a dissertation on domain-specific failures of automatic evaluation in natural language generation.

The post itself is brief and celebratory, but the referenced thesis topic is a real AI research contribution: it focuses on where automatic evaluation breaks down when NLG systems are judged in specific domains, rather than by generic metrics alone.

Original post →

More from Research

Research channel →