NVIDIA at IFA 2026: 1.9x faster local inference, PAIR router, RTX Spark PCs in October

nordicinst · x · 2026-09-04

At IFA 2026, NVIDIA announced up to 1.9x faster local inference via new llama.cpp and vLLM optimizations (also in LM Studio and Ollama), NVIDIA PAIR — a Personal AI Router that spreads inference across PCs on a home network — and compact RTX Spark Windows PCs from Lenovo and Acer arriving in October. New local-capable models include 30B Nemotron 3.5 Lightning, Z.ai's GLM-5.3-Flash MoE, and Qwen3.8-Flash-Next / Qwen3.8-27B open-weight models for DGX Spark and DGX Station.

Original post →

More from Embodied

Embodied channel →