Nvidia's race to manufacture Groq chips and make them available to customers highlights the growing importance in AI of low-latency inference.

NVIDIA Groq 3 LPX is the interactive AI inference accelerator for the NVIDIA Vera Rubin platform. At the core of the platform is NVIDIA Vera Rubin NVL72, the most versatile…

Gemma 4 31B performance tests offer a best-case scenario for next-gen dataflow accelerators