Leaky Language Models show token timing can expose architecture and optimizations

chaumian · x · 2026-07-24

Leaky Language Models show timing can expose architecture and inference optimizations

A new paper titled Leaky Language Models argues that per-token timing can leak enough information to infer a model’s architecture and some inference optimizations.

Original post →

More from Safety

Safety channel →