IBM has signed a $240 million multi-year agreement with Together AI to deploy a “large cluster” of Nvidia HGX B300 systems on IBM Cloud.
Interconnected by Nvidia Spectrum-X Ethernet technology, the deployment is the first large-scale cluster built for inference on IBM Cloud using HGX B300 systems.
Together AI said it selected IBM and Nvidia because of their ability to “deliver GPU capacity at the pace required for rapid AI scaling and lowest token cost.” The cluster is expected to be available from Q1 2027.
"Enterprises want the performance of the best frontier models without the closed-model price tag, and that only works if the infrastructure underneath is fast and reliable at scale," said Vipul Ved Prakash, CEO at Together AI. "Working alongside IBM with Nvidia gives us that foundation. This cluster lets us bring production-grade inference to more companies, faster, and it's a big step in our push to make open-source AI the obvious choice for enterprises."
Alan Peacock, general manager of IBM Cloud, added: "Enterprises are in a race to adopt agentic AI at scale to drive real business outcomes. IBM and NVIDIA are delivering scalable, economical, enterprise-grade AI infrastructure that can help Together AI accelerate innovation for the next generation of AI infrastructure."
Together AI was founded in 2022. The self-described AI acceleration cloud company both leases chips from other cloud providers and re-leases them to developers, and is buying its own servers to rent from its data centers. Details about Together AI's data centers are sparse, but the company's documentation says they are located in North America.
In July 2026, the AI cloud company raised $800m in its Series C funding round, which saw the company achieving an $8.3bn post-money valuation. The month prior, Together AI signed a multi-year cloud capacity agreement with AI compute provider Rumble Inc. Under the contract, Rumble will also provide Together AI with Nvidia HGX B300 systems as dedicated cloud capacity.
Comments