Bilevel optimization densifies scarce labels to fix OOD molecular property prediction (ICML WS 2025)
CatAstro_Piyush · x · 2026-09-26
Accepted at GenBio ICML Workshop 2025 (arXiv:2506.11877), this work by Jina Kim, Jeffrey Willette, Bruno Andreis, and Sung Ju Hwang tackles a core drug-discovery problem: molecular prediction models generalize poorly to out-of-distribution compounds, and labeled data is scarce because experimental validation is costly. The authors propose a bilevel optimization approach that leverages unlabeled data to interpolate between in-distribution and OOD samples, teaching the model to generalize beyond the training set. They report significant gains on real-world datasets with heavy covariate shift, supported by t-SNE visualizations.
More from Research
- WetLabs Benchmark tests robots on 9 wet-lab tasks: can machines accelerate science? — ericjang11 · 2026-09-26
- Medmarks v1.0 lands NeurIPS track: medical LLM benchmark now covers 30 suites, 61 models — iScienceLuvr · 2026-09-26
- AI-driven lab finds iridium-free palladium catalyst InMnPdOx stable for 1,000 hours — CatAstro_Piyush · 2026-09-26
- RGBD20K: new RGB-D segmentation benchmark with 20K image pairs and 160 categories — UNT · 2026-09-26
- Formally verified Ethereum consensus proposal moves toward 4-8x faster finality — DavideCrapis · 2026-09-26
- Fewer than 10% of preregistered studies include multiple-hypothesis corrections — RexDouglass · 2026-09-26