IBM and Together AI announced a $240 million multi-year collaboration to deploy the first large-scale NVIDIA HGX B300 in
IBM and Together AI announced a strategic collaboration to deliver AI infrastructure on IBM Cloud. Under a multi-year $240 million agreement, IBM will deploy a large cluster of NVIDIA HGX B300 systems with expected availability in Q1 2027. Together AI will use this cluster to provide open-source model inference capabilities to enterprises. This marks the first dedicated, large-scale cluster built for inference on IBM Cloud using HGX B300 systems and NVIDIA Spectrum-X Ethernet networking. According to NVIDIA, the deployment is built to deliver 30x more AI factory output compared to prior generations.
The collaboration aims to enable Together AI to deliver improved performance and token economics to enterprises scaling their AI deployments. Together AI, which operates on the principle that open-source models are essential for AI's future, recently raised an $800 million Series C financing round at an $8.3 billion valuation. The platform spans capabilities across inference, training, fine tuning, and agentic workflows. Together AI reports it now serves 400 trillion tokens monthly. The company selected IBM and NVIDIA based on their innovative product roadmaps and ability to deliver GPU capacity at the pace required for rapid AI scaling and lowest token cost.
Vipul Ved Prakash, CEO at Together AI, stated that enterprises want the performance of frontier models without closed-model price tags, and that requires fast, reliable infrastructure at scale. He noted that the cluster enables production-grade inference for more companies and represents a significant step in making open-source AI the obvious choice for enterprises.
Alan Peacock, General Manager of IBM Cloud, emphasized that enterprises are racing to adopt agentic AI at scale and that IBM and NVIDIA are delivering the scalable, economical, enterprise-grade AI infrastructure to support this transition. Dion Harris, Senior Director of HPC and AI Infrastructure Solutions at NVIDIA, described AI factories as becoming essential enterprise infrastructure comparable to electricity and telecommunications, turning compute and data into intelligence with performance and efficiency at the required scale.
This collaboration builds on the broader IBM and NVIDIA partnership, which recently announced progress across GPU-native data analytics, unstructured data extraction, on-premises and cloud infrastructure, and consulting services designed to help organizations operationalize AI at scale.