SMDD-Bench tests whether LLM agents can budget through real small-molecule design
niloofar_mire · x · 2026-07-23
Carnegie Mellon researchers say their AI-for-science work, including SMDD-Bench, is one of three projects selected for U.S. Department of Energy Genesis Mission awards.
The referenced benchmark, SMDD-Bench, evaluates long-horizon agentic small-molecule design. It includes 502 guaranteed-solvable tasks across five real drug-design workflows — pharmacophore identification, scaffold hopping, lead optimization, fragment assembly, and interaction point discovery — and gives agents a Python sandbox plus strict limits on Boltz2 and ADMET-AI calls. The goal is to test whether frontier LLM agents can plan and budget through multi-turn medicinal chemistry tasks instead of handling toy single-step prompts.
Related event: CMU Receives Three DOE Awards for AI-Driven Molecular Design(2 posts)→
More from Companies & People
- Peter Diamandis says enterprise in-house models could reset top lab valuations — PeterDiamandis · 2026-07-23
- White House Accuses Moonshot AI of Distilling Anthropic Models; Jensen Huang Pushes Back — TheTuringPost · 2026-07-23
- LangChain promotes Interrupt London as an agent conference with workshops — LangChain · 2026-07-23
- Google Cloud backlog hits $462B as GenAI product revenue rises 800% YoY — Beth_Kindig · 2026-07-23
- Liang Wenfeng transcript highlights a four-hour investor meeting — zephyr_z9 · 2026-07-23
- Anatomy of the Twitter Hack: Fake Journalists Weaponizing Calendly Links — giffmana · 2026-07-23