NVIDIA Releases Open-Source 30B Nemotron Model Optimized for AI Agents
DeryaTR_ · x · 2026-08-12
Following Meta's releases, NVIDIA has introduced Nemotron 3.5 Lightning, an open-weights 30B parameter Mixture-of-Experts (MoE) model with 3B active parameters.
Designed for always-on agents, the model is built to handle high-volume, specialized tasks rapidly, delivering up to 4x the output speed of similar-sized models. It is now available on Hugging Face and can be post-trained using the NVIDIA NeMo framework for domain-specific applications across cybersecurity, coding, legal, and energy tasks.
Related event: NVIDIA Open-Sources Nemotron Model and Switchyard Router(45 posts)→
More from Infra
- AMD FastFlowLM 1.0 Released and Integrated into the ROCm Ecosystem — AnushElangovan · 2026-08-12
- Oracle's AI Infrastructure Push Turns Cash Flow Negative, Plans Major Layoffs — rohanpaul_ai · 2026-08-12
- Lumentum Earnings Quell Rumors, Confirms Accelerated Nvidia CPO Demand — zephyr_z9 · 2026-08-12
- Hyperscalers Still Rely on 2017's V100s: AI Compute Lifespan Reaches 9 Years — BenBajarin · 2026-08-12
- TensorScale Unveils Fastest Video Inference, Claims 10x Speedup for MiniMax H3 — Scobleizer · 2026-08-12
- vLLM and NVIDIA Co-host Meetup on Scaling LLM Inference Efficiency — vllm_project · 2026-08-12