NVIDIA Q&A: Nemotron 3.5 Lightning sub-agents, vLLM or Ollama?
NVIDIA Developer · youtube · 2026-10-01
NVIDIA Developer releases a Q&A video on Nemotron 3.5 Lightning, covering how to build sub-agents with the model and choosing between vLLM and Ollama for local deployment.
More from Infra
- Broadcom to lend Anthropic up to $42B for AI chips, eyeing top customer slot by 2027 — rohanpaul_ai · 2026-10-02
- Meta paper: only 50-60% of recommendation training time actually trained before optimizations — _reachsumit · 2026-10-02
- Dev open-sources GPT-2-tools to run original 1.5B GPT-2 XL locally on CPU — MikePFrank · 2026-10-02
- CoreWeave launches serverless GPUs: hourly-billed, no contract, private preview — altryne · 2026-10-02
- RTX 5090 + 5070 Ti workstation: doubling VRAM wasn't worth it for local LLMs — Lordofwhut · 2026-10-02
- 6x BC-250 mining board cluster runs local LLMs at 100k context, 28 tok/s — Ok-Breadfruit-3523 · 2026-10-02