Traditional cloud infrastructure was designed for predictable, general-purpose workloads and standard interfaces. Agentic AI factories connect diverse users, agents, applications, data sources, and storage systems to massively accelerated compute at multi-terabit bandwidth per server, making dedicated DPU processing essential for line-rate networking, storage, and security. NVIDIA is introducing Scale-In network infrastructure, the fifth pillar of NVIDIA AI networking, bringing purpose-built acceleration to secure, manage, and operate agentic AI factories.

Scale-In evolves north-south networks into a coordinated infrastructure domain for the AI factory. Powered by NVIDIA BlueField-4 and NVIDIA DOCA and connected over NVIDIA Spectrum-X Ethernet, Scale-In accelerates the services that secure the full AI stack and move application, data, and storage traffic across the AI factory. Dedicated, host-independent processing keeps these infrastructure services off host CPUs, helping prevent security, data access, and operations from becoming bottlenecks as AI compute scales. The result is a more secure and efficient shared AI infrastructure with consistent access to AI services and data as demand grows.