NVIDIA publishes PAIR, an Apache-2.0 router that spreads local inference across machines on one network
- Sources: primary, product page, discussion
- Summary: NVIDIA published the Personal AI Router under Apache-2.0. It distributes independent inference requests across machines on a single network, sending each request to one node.
- Why it matters: PAIR routes each independent request to one node and states it does not pool GPU memory, shard a model across machines, or split an in-flight request, so it adds concurrency rather than capacity for larger models.