NVIDIA releases Nemotron 3.5 Lightning: 30B LatentMoE, 3B active, 1M context
huggingface · x · 2026-08-11
NVIDIA unveiled Nemotron 3.5 Lightning for more efficient agents. It features a 30B LatentMoE with 3B active parameters, supports NVFP4 and BF16, up to 1M context, MTP, DFlash, DSpark, tool use, and multilingual. Integrated with NousResearch's Hermes Agent and runs via Microsoft Foundry.
Related event: NVIDIA Open-Sources Nemotron Model and Switchyard Router(27 posts)→
More from Models
- OpenRouter Data: Reasoning Model Token Share Exceeds 60% — Beth_Kindig · 2026-08-11
- Nvidia Releases Nemotron 3.5 Lightning Open 30B MoE Model — dr_alphalyrae · 2026-08-11
- Nemotron 3.5 Specs Revealed: 3.6B Active Params Matches Larger Models — thisguyknowsai · 2026-08-11
- LlamaIndex Launches ExtractBench: A New Benchmark for Enterprise Document Extraction — llama_index · 2026-08-11
- Benchmarking Meta's New Muse Glimmer: A 30B Model Optimized for Agents — altryne · 2026-08-11
- NVIDIA Updates 30B Hybrid Model: Distillation Delivers 70% Intelligence Boost — PavloMolchanov · 2026-08-11