Bio ML critique: many published models score within ~10% of a naive linear baseline
iskander · x · 2026-09-08
iskander (George Ho) observes that many published bio ML models score within roughly 10% of a naive linear baseline.
The short note points to a familiar pitfall in bio ML papers: models are benchmarked against deep learning baselines rather than simple linear regressions, making seemingly sophisticated models look better than they are.
More from Research
- Cheating is rampant in modern AI benchmarks, says Terminal-Bench contributor — xeophon · 2026-09-08
- Looped Transformer hype: Nanbeige4.2-3B beats 12B models on agent benchmarks — alexcovo_eth · 2026-09-08
- 100,000-woman randomized trial offers lessons on judging medical AI by patient outcomes — EricTopol · 2026-09-08
- Researchers flag massive reporting bias in AI math capabilities: failures go untracked — RexDouglass · 2026-09-08
- Fields Medalist Voevodsky on Why He Started Verifying All His Proofs in Coq — RexDouglass · 2026-09-08
- PlaidQ: 0.7B continuous diffusion LM distilled to one step for code generation — AlexanderTong7 · 2026-09-08