NVIDIA porta in produzione Groq 3 LPX, acceleratore dedicato all'inferenza AI a bassissima latenza, integrato nella piattaforma Vera Rubin. Il sistema raggiunge 3.400 token al secondo nei test Artificial Analysis e sar� adottato per primo da Nebius. L'obiettivo � accelerare gli agenti AI senza sostituire le GPU tradizionali.

Gemma 4 31B performance tests offer a best-case scenario for next-gen dataflow accelerators

NVIDIA Groq 3 LPX is the interactive AI inference accelerator for the NVIDIA Vera Rubin platform. At the core of the platform is NVIDIA Vera Rubin NVL72, the most versatile…