Bergson: open-source library unifies data attribution methods to study LLM generalization and misalignment

zetalyrae · x · 2026-10-05

A new paper from luciaquirke's team tackles a key question: how do LLMs generalize, and why do they sometimes become misaligned? Their angle: the training data.

The team released Bergson, an open-source library that brings together data attribution methods across a wide range of compute budgets, letting researchers trace model behavior back to training samples and study generalization and alignment issues with whatever resources they have.

Original post →

More from Research

Research channel →