NVIDIA's open-source PAIR beta routes local AI inference across PCs on your network
Codeblix_Ltd · reddit · 2026-09-04
NVIDIA released PAIR, a free open-source beta tool that discovers compatible PCs on a local network and routes independent inference requests to whichever machine has spare capacity.
- Works with Ollama and LM Studio; supports Windows, macOS, and Linux
- Hardware: RTX 20-series and newer, RTX PRO workstation GPUs, DGX Spark, and Apple M4 or newer
- Aimed at local agent workflows that split tasks into smaller jobs across machines
- Related IFA updates simplify local model setup for Hermes Agent, OpenClaw, and Perplexity Portable Computer, and vendor-reported up to 1.9x llama.cpp throughput on an RTX 5090
- Open question: whether multi-PC routing helps real workflows without complicating setup and privacy
More from Infra
- prime-rl ships blazingly fast weight transfer for RL training pipelines — TheZachMueller · 2026-09-04
- OpenAI adds AWS as partner, commenters say the Microsoft marriage is officially over — zephyr_z9 · 2026-09-04
- Cloudflare frees 100TB of memory, shrinks DNS cache records from 953 to 420 bytes — shashib · 2026-09-04
- Grok outage traced to Memphis compute center failure, Musk says corrective action taken — NicoVerderosa · 2026-09-04
- NVIDIA Parabricks HaplotypeCaller lands in nf-core/sarek for GPU-accelerated germline calling — AllThingsApx · 2026-09-04
- Jane Street signs ~$13B five-year AI cloud deal with Crusoe, Bloomberg reports — IanAndrewsDC · 2026-09-04