John Schulman: Forensics for Model Training Data and Backdoors Understudied
johnschulman2 · x · 2026-07-27
John Schulman points out that while model weights alone don't reveal training data, methods, or potential backdoors, it's possible to perform extensive forensics to figure these things out. He notes that this forensic problem is currently under-studied.
Related event: Experts Discuss Forensics for Open-Source Models(3 posts)→
More from Research
- Minerva paper details: ncRNA arrays template short DNA hairpins — BrianHie · 2026-09-23
- Genome language models uncover new class of reverse-transcriptase mechanisms — BrianHie · 2026-09-23
- Mathematician shares a cheap 4-step heuristic for hyperparameter tuning — dejanseo · 2026-09-23
- Burkov skew AI hype: 'deterministic LLMs' and 'first agents' are old tricks rebranded — burkov · 2026-09-23
- Continuous diffusion beats discrete on random k-SAT, proposed as standard benchmark — ArashVahdat · 2026-09-23
- Grady Booch: Contemporary AI Still Lacks Abductive Reasoning, Just 'Next-Token Prediction' — Grady_Booch · 2026-09-23