Equinix, Together AI, NVIDIA는 기업에 안전하고 저지연의 NVIDIA 컴퓨팅 및 200개 이상의 오픈 모델 접근을 제공하기 위해 Inference Exchange를 출시했다.
Data center operator Equinix is partnering with Nvidia and Together AI to develop a distributed AI inference platform called the Inference Exchange. Built on Nvidia’s Enterprise Reference Architecture and infrastructure, combined with Together AI’s inference software, the solution will be delivered through Equinix’s global network of facilities.
Equinix operates more than 280 data centers across 77 metros, provides 230 cloud on-ramps, and interconnects over 10,500 businesses on its exchange. Through Equinix Fabric, the Inference Exchange will link directly to public clouds, telecommunications networks, and third-party AI providers.
“AI is transforming enterprise technology at extraordinary speed, and the infrastructure decisions enterprises make today will define their competitive position for years to come,” said Adaire Fox-Martin, CEO and president, Equinix. “Equinix is uniquely positioned to deliver what this moment demands based on our nearly three decades building the trusted exchange where the world's enterprises run, connect, and orchestrate their most critical workloads.”
“Our longtime relationship with Nvidia delivers the accelerated computing foundation at the heart of modern AI, while Together AI's commitment to open ecosystems gives enterprises the flexibility to scale on their terms. Equinix Inference Exchange will enable architectures that are neutral by design, open by default, and engineered for exceptional performance.”
“Together AI was built on the conviction that open, accessible AI is what will define the industry moving forward, because enterprises shouldn't have to choose between model performance and operational flexibility,” added Vipul Ved Prakash, co-founder and CEO of Together AI. “What we are building with Equinix and Nvidia proves that model choice and performance are not trade-offs. They are the foundation of enterprise AI done right.”
The offering targets several key use cases, including metro Edge inference for organizations requiring low latency for AI workloads, open model migration for enterprises looking to move from closed proprietary models to open-source alternatives using Together AI's platform, and sovereign AI workloads that need to run in specific locations.
The Inference Exchange will be available from Q1 2027.
AI cloud company Together AI raised $800m in a Series C funding round in July 2026. The company is a known customer of Rum Group and is planning to deploy hardware at an L&T data center, while Hypertec and 5C are aiming to roll out 2GW of capacity for the company.