IBM And Together AI Sign $240 Million Multi-Year Deal For Large-Scale NVIDIA B300 AI Inference Cluster

By Amit Chowdhry ● Today at 4:04 PM

IBM has signed a multi-year $240 million agreement with Together AI to deploy a large-scale artificial intelligence inference cluster on IBM Cloud using NVIDIA HGX B300 systems. The cluster is expected to become available in the first quarter of 2027 and will help Together AI provide open-source model inference for enterprise customers.

IBM said the deployment will be the first dedicated, large-scale inference cluster on IBM Cloud built with NVIDIA HGX B300 systems and NVIDIA Spectrum-X Ethernet networking.

According to NVIDIA, the infrastructure is designed to deliver up to 30 times more AI factory output than prior generations.

Together AI plans to use the cluster to improve inference performance and token economics as enterprises move larger AI workloads into production.

The company is focused on open-source artificial intelligence and operates an AI Native Cloud platform spanning inference, training, fine-tuning and agentic workflows.

Together AI recently raised an $800 million Series C round at an $8.3 billion valuation.

The company said its inference platform is now serving approximately 400 trillion tokens per month, reflecting growing demand for open-source models in production environments.

Together AI selected IBM and NVIDIA as infrastructure partners based on their product roadmaps and ability to provide GPU capacity at the scale required for rapidly expanding AI workloads.

The companies expect the deployment to give developers and enterprises access to large-scale open-source inference through IBM’s enterprise cloud infrastructure.

IBM is positioning the collaboration around growing demand for agentic AI and increasingly compute-intensive enterprise applications.

The infrastructure will combine IBM Cloud, NVIDIA GPU systems, Spectrum-X networking and Together AI’s inference software.

For IBM, the agreement also expands its broader relationship with NVIDIA across AI infrastructure and software.

The companies have been working together across GPU-native data analytics, unstructured data extraction, cloud and on-premises infrastructure, and consulting services designed to help enterprises deploy AI.

The latest deployment is intended to provide Together AI with infrastructure capable of supporting real-time AI services while expanding its presence among large enterprises.

KEY QUOTES:

“Enterprises want the performance of the best frontier models without the closed-model price tag, and that only works if the infrastructure underneath is fast and reliable at scale. Working alongside IBM with NVIDIA gives us that foundation. This cluster lets us bring production-grade inference to more companies, faster, and it’s a big step in our push to make open-source AI the obvious choice for enterprises.”

Vipul Ved Prakash, CEO Of Together AI

“Enterprises are in a race to adopt agentic AI at scale to drive real business outcomes. IBM and NVIDIA are delivering scalable, economical, enterprise-grade AI infrastructure that can help Together AI accelerate innovation for the next generation of AI infrastructure.”

Alan Peacock, General Manager Of IBM Cloud

“AI factories are becoming essential enterprise infrastructure, like electricity and telecommunications, turning compute and data into intelligence. With NVIDIA HGX B300 systems and NVIDIA Spectrum-X Ethernet networking on IBM Cloud, IBM and Together AI will deliver an accelerated computing platform to help enterprises deploy open-source AI with the performance, efficiency and scale required for real-time AI services.”

Dion Harris, Senior Director Of HPC And AI Infrastructure Solutions At NVIDIA

Exit mobile version