NVIDIA has released PAIR, a new software tool that transforms idle home gaming PCs into a local AI cluster. This update matters because it allows users to leverage existing hardware for artificial intelligence tasks instead of paying for expensive cloud services. The release offers a practical way to reduce monthly expenses for users who frequently run large language models.

Aggregating idle home GPUs cuts cloud API costs by up to $1,200 monthly
The software is designed to aggregate inference compute from multiple devices on a local network. It supports Windows, Linux, and macOS, ensuring broad compatibility with most modern home setups. PAIR proxies compatible Ollama and LM Studio interfaces, making it easy to integrate with existing workflows.
PAIR uses mDNS to automatically discover devices and pairs them using a secure 6-digit code. All communication between devices is encrypted through MTLS to ensure security. The tool routes inference jobs to available nodes while actively avoiding devices that are currently running games or heavy workloads.
NVIDIA claims the software can reduce multi-agent workflow completion times by approximately 40% compared to single-device execution. In tests using Qwen3.6 35B A3B on RTX 5090 GPUs, a two-device cluster completed tasks in 3 minutes and 48 seconds versus 6 minutes and 18 seconds on a single unit. The company estimates this efficiency could save average US homes roughly $1,200 per month in cloud API costs.
The PAIR release is currently in BETA and is available globally starting in September 2026. NVIDIA has open-sourced the software under the Apache 2.0 license. The company plans to work toward a full release after the beta period concludes.



Discussion
0 comments
Log in to join the thread with a thoughtful take, question, or correction.