"Disaggregated Inference," promises better utilization, lower costs, and faster AI responses. Major players like NVIDIA/Groq and AMD/Cerebras will vie for the prize.

"Disaggregated Inference," promises better utilization, lower costs, and faster AI responses. Major players like NVIDIA/Groq and AMD/Cerebras will vie for the prize.

Cerebras CMO Julie Choi says the AMD Helios and Cerebras Wafer-Scale Engine partnership delivers 5X higher throughput for disaggregated AI inference.