Your Model May Be Learning Who Generated the Data — A Robotic Surgery Case

bravo_abad · x · 2026-09-30

AI-for-science writer Jorge Bravo Abad highlights an overlooked pitfall: models can pick up the habits of whoever generated the data. A model trained on scientific measurements may learn the operator's habits; if those habits also predict the label, a random train–test split makes the model look more useful than it is.

Ueki and colleagues provide a striking example in robotic surgery: analyzing 98 procedures by 16 surgeons with 3D hand-tracking data, they found exactly this kind of operator leakage.

The author explores these ideas further on Discovery at Scale, where he also writes a weekly AI for Science briefing.

Related event: Study Finds Models Can Learn the Habits of Data Generators(2 posts)→

Original post →

More from Research

Research channel →