IBM has signed a $240 million multi-year agreement with Together AI to deploy a “large cluster” of Nvidia HGX B300 systems on IBM Cloud.
Interconnected by Nvidia Spectrum-X Ethernet technology, the deployment is the first large-scale cluster built for inference on IBM Cloud using HGX B300 systems.
– Getty Images
Together AI said it selected IBM and Nvidia because of their ability to “deliver GPU capacity at the pace required for rapid AI scaling and lowest token cost.” The cluster is expected to be available from Q1 2027.
"Enterprises want the performance of the best frontier models without the closed-model price tag, and that only works if the infrastructure underneath is fast and reliable at scale," said Vipul Ved Prakash, CEO at Together AI. "Working alongside IBM with Nvidia gives us that foundation. This cluster lets us bring production-grade inference to more companies, faster, and it's a big step in our push to make open-source AI the obvious choice for enterprises."






