News
Nvidia Shipped a Free Tool That Pools Every Machine on Your LAN Into One Inference Cluster
At IFA 2026 Nvidia released PAIR (Personal AI Router) in beta, a free open-source tool that distributes inference requests across whichever PCs on a local network have spare capacity, letting agentic workflows run in parallel. It supports GeForce RTX 20 Series and newer, RTX PRO workstation GPUs from Turing on, DGX Spark, and Apple M4 or newer silicon, and works with Ollama and LM Studio on Windows, macOS and Linux. Nvidia also cited 1.9x llama.cpp throughput on RTX 5090 and 1.2x-1.4x vLLM gains, plus one-click setup for Nous Research's Hermes Agent and OpenClaw, with RTX Spark PCs from Lenovo and Acer arriving in October.
Source
↳ Follow the thread