OpenAI's custom Jalapeño inference chip outperformed Nvidia's GB300 processors by up to 3.6x in latency tests, with a planned rollout beginning

128 chips, 1.7 exaFLOPS, and 27 TB of HBM give Altman and crew a leg up over Blackwell, and maybe even Rubin

OpenAI still isn’t giving up Nvidia chips, though.