IBM and Together AI Sign Multi-Year Agreement to Scale Open-Source AI Inference with NVIDIA AI Infrastructure on IBM Cloud
IBM and Together AI will deploy a $240M NVIDIA HGX B300-based inference cluster on IBM Cloud by Q1 2027 to scale open-source AI workloads for enterprises.
IBM and Together AI announced a multi-year $240 million agreement to deploy a large-scale inference cluster on IBM Cloud using NVIDIA HGX B300 systems, with expected availability in the first quarter of 2027. The cluster will be the first dedicated large-scale inference deployment on IBM Cloud using HGX B300 systems and NVIDIA Spectrum-X Ethernet networking. Together AI will use the infrastructure to provide open-source model inference services to enterprises seeking scalable AI solutions.
According to NVIDIA, the deployment is designed to deliver up to 30 times more AI factory output compared to prior generations, aiming to improve performance and token economics for enterprise AI workloads. Together AI, which operates on an open-source model philosophy, recently raised $800 million in a Series C round at an $8.3 billion valuation to expand its AI Native Cloud platform. The company reports processing 400 trillion tokens monthly through its inference product.
Together AI selected IBM and NVIDIA for their product roadmaps and ability to deliver GPU capacity at the required pace for rapid AI scaling and cost efficiency. The collaboration builds on IBM’s enterprise-grade cloud capabilities to support Together AI’s expansion into enterprise markets while promoting accessibility of open-source AI technologies. Vipul Ved Prakash, CEO of Together AI, emphasized the need for fast, reliable infrastructure to enable cost-effective open-source AI adoption by enterprises.
IBM Cloud General Manager Alan Peacock highlighted the infrastructure’s role in helping enterprises adopt agentic AI at scale to achieve business outcomes. NVIDIA’s Dion Harris described AI factories as essential enterprise infrastructure, noting the deployment will provide the performance, efficiency, and scale required for real-time AI services. The collaboration reflects a broader IBM-NVIDIA partnership advancing AI infrastructure and software for enterprise and startup clients.