The new lightweight open model and routing library delivers greater control over AI, data and workflows across edge devices, PCs, workstations, data centers and the cloud.

Learn how NVIDIA NeMo Switchyard routes AI agent workloads across models using tuning-free and tunable routers that balance model capability, cost, and latency.

Nvidia's Nemotron 3.5 Lightning pairs with its NeMo Switchyard router, which reassigns models mid-task and cuts task costs to a third in Nvidia's own tests.