The NVIDIA PAIR beta connects AI app and agent workflows to a single local endpoint for routing inference across NVIDIA DGX Spark™, Windows systems with RTX™, and macOS devices. This helps you maximize local compute while keeping prompts, files, and agent context private.
Features
Bring together the RTX, DGX™, and Mac systems already on your network in minutes. NVIDIA PAIR discovers compatible local machines and helps them work as one personal AI inference cluster with no special cables, racks, or complex cluster setup required.
Keep local AI workflows moving when tasks stack up. NVIDIA PAIR routes AI inference requests across available local nodes, helping busy AI workflows tap into idle compute regardless of the node’s operating system.
Run NVIDIA PAIR alongside familiar local inference backends on Windows, Linux, and macOS. At launch, NVIDIA PAIR supports Ollama and LM Studio, giving apps a consistent, single endpoint while intelligently proxying requests to available local compute.
Run AI workflows at home with data that stays on your network. NVIDIA PAIR is built for private local inference, helping you use prompts, files, and agent context without sending them to the cloud.
NVIDIA PAIR works without special cables or racks. And setting up is easy.
Platform
Windows 11, DGX OS, Ubuntu 14.04, macOS Tahoe
Language
English
GPU
All GeForce RTX GPUs (20 Series and newer), DGX Spark/GB10, Mac M4 or newer
RAM
8 GB RAM or higher
Disc Space
Recommended: 20 GB or higher
Internet
None required for operation
Required for model download
NVIDIA Personal AI Router (PAIR) is software that connects compatible macOS, Windows, and Linux systems with NVIDIA RTX™GPUs and DGX Spark systems into a personal home AI cluster. PAIR distributes local AI inference workloads across available devices while keeping prompts, files, and agent context on the user’s home network.
A personal home AI cluster connects multiple compatible devices on the same local network, allowing them to share available computing resources for local AI inference. PAIR enables Macs, NVIDIA RTX systems, and DGX Spark systems to provide more capacity for local AI applications and agents. The devices remain separate systems that handle parallel tasks; PAIR doesn’t combine them into one virtual GPU.
Download PAIR, install it on supported systems, pair the devices on the same local network, and add them to the cluster. Once connected, supported applications and agents can send inference requests through the PAIR local endpoint.
NVIDIA PAIR requires compatible macOS, Windows, or Linux systems and supported hardware. Requirements may include specific operating system versions, NVIDIA RTX GPUs, GPU memory, system memory, processors, storage, and local network connectivity. See the System Requirements section on the NVIDIA PAIR landing page for current details.
PAIR is designed for private local inference. Prompts, files, and agent context remain on the user’s local network instead of being sent to a cloud inference service.