John Schulman breaks down three ways AI firms train on user data

anshulkundaje · x · 2026-09-15

John Schulman explains that 'training on user data' spans very different practices with distinct privacy/IP risks, and AI companies rarely disclose which they use:

Sarah Hooker adds there are synthetic-data techniques generating distributionally equivalent data while preserving privacy.

Original post →

More from Safety

Safety channel →