NVIDIA open-sources BioNeMo Inference Runtime, boosting protein structure prediction throughput 2.9x
AllThingsApx · x · 2026-09-11
NVIDIA has opened public beta of BioNeMo Inference Runtime, a PyTorch-native open-source library that accelerates biomolecular structure prediction on NVIDIA GPUs.
- Models accelerated: Boltz-2, OpenFold2, and Protenix v2 via dedicated kernels and CUDA Graphs, with Ray-driven GPU replicas for scaling.
- Benchmark: On 1,000 human dimer targets across 8xH100 GPUs, accelerated Boltz-2 folded 58.5K residues per GPU-hour vs 20.2K for a torch-compiled open-source baseline — a 2.90x residue-normalized throughput gain; energy to fold 1M targets drops from 35 MWh to 11 MWh.
- Three optimization layers: kernel selection, CUDA Graph module optimization, and Ray pipeline scaling.
- Partners include SandboxAQ, Xaira Therapeutics, Latent Labs, and the UW Institute for Protein Design.
More from Infra
- SageMaker prefix-aware routing cuts P50 TTFT by up to 77% via warm KV caches — AWS ML Blog · 2026-09-11
- Jensen Huang explains why Nvidia will grow 70% next year, denies circular deals — TechCrunch AI · 2026-09-11
- Microsoft plans to grow data center capacity from ~12 GW to 38+ GW by 2032, per Bloomberg — BenBajarin · 2026-09-11
- Oracle beats estimates with $19.3B Q1 revenue as AI cloud demand surges — Polymarket · 2026-09-11
- As AI agents grow capable, more compute shifts from GPUs to CPUs — AccBalanced · 2026-09-11
- 2.78T-param Kimi K3 runs inference on a single CPU in 8.24 GB RAM, no GPU or BLAS — udmrzn · 2026-09-11