Nvidia’s LPX racks for AI inference accelerators have entered full production, the company has confirmed.
Unveiled at GTC back in March, the rack-scale platform came about following Nvidia’s acqui-hire of the eponymous startup. LPX is liquid-cooled and houses some 256 Groq 3 language processing units (LPUs) interconnected through some 640 terabits per second (Tbps) of scale-up bandwidth.
A glimpse inside the LPX from GTC 2026 – Sebastian Moss
Inside the LPX rack itself are BlueField-4 data processing units, Vera CPU racks, and STX storage servers all tied together with the recently debuted Spectrum-6 Ethernet networking tech.
The platform is not a replacement for Nvidia’s flagship NVL72 platform, but rather a complementary add-on for operators wanting to power ultra-low-latency AI inference workloads.






