NVIDIA BioNeMo Inference Runtime Enters Public Beta, Speeds Protein Models Like Boltz-2 on GPUs
AllThingsApx · x · 2026-09-11
NVIDIA has released BioNeMo Inference Runtime in public beta — an open, PyTorch-native library that accelerates biomolecular inference on NVIDIA GPUs for drug discovery and protein design.
Key points:
- Optimizes structure prediction models including Boltz-2, OpenFold2, and Protenix v2 with specialized CUDA kernels and CUDA Graphs where supported
- Scales with Ray-powered GPU replicas
- Built with collaborators including Apheris, Xaira Therapeutics, SandboxAQ, Latent Labs, and the Institute for Protein Design at UW
More from Infra
- Microsoft plans to grow data center capacity from ~12 GW to 38+ GW by 2032, per Bloomberg — BenBajarin · 2026-09-11
- Oracle beats estimates with $19.3B Q1 revenue as AI cloud demand surges — Polymarket · 2026-09-11
- As AI agents grow capable, more compute shifts from GPUs to CPUs — AccBalanced · 2026-09-11
- 2.78T-param Kimi K3 runs inference on a single CPU in 8.24 GB RAM, no GPU or BLAS — udmrzn · 2026-09-11
- Cherenkov engine hits 8-22 tok/s Qwen3.8-Flash-Next on a 32GB M4 MacBook Air — alfredr · 2026-09-11
- Keep the Claude Desktop Workflow, Swap in Local Models via Ollama for Privacy — Technovangelist · 2026-09-11