Friday, September 11, 2026
AI 인프라 · 뉴스 & 분석
데이터센터리포트
데이터센터 · 리포트

Equinix는 NVIDIA 및 Together AI와 파트너십을 맺어 기업용 콜로케이션 시설에 최적화된 AI 추론 랙을 구축한다.

비하이퍼스케일 기업의 접근 가능한 저지연 추론 용량을 확대하여, 기존 학습 클러스터 중심의 테넌트 수요를 다각화합니다.
업계 전문지Slicast · September 5, 2026 · 미국 · 출처: iTWire
중요도 82

Equinix Inference Exchange combines NVIDIA Enterprise Reference Architectures, Together AI’s inference platform, and Equinix’s global infrastructure to optimize deployment speed, flexibility, and cost efficiency. Equinix, the world’s digital infrastructure company, has announced a significant expansion of its longstanding collaboration with NVIDIA to deliver Equinix Inference Exchange, a distributed AI inference program for global enterprises, alongside a new partnership with Together AI.

As AI scales across models, providers, and geographies, the location of inference has become a strategic imperative that dictates performance, cost, and governance. Equinix Inference Exchange will give enterprises a faster path from AI experimentation to production, providing secure, low-latency connectivity to the data, users, and ecosystem they depend on.

This collaboration unites NVIDIA’s validated Enterprise Reference Architectures with Together AI’s inference platform, which supports more than 200 open-source models. Delivered through Equinix’s global data centers, the solution will provide seamless connectivity to clouds, networks, and AI providers via Equinix Fabric.

“AI is transforming enterprise technology at extraordinary speed, and the infrastructure decisions enterprises make today will define their competitive position for years to come. Equinix is uniquely positioned to deliver what this moment demands based on our nearly three decades building the trusted exchange where the world’s enterprises run, connect and orchestrate their most critical workloads,” said Adaire Fox-Martin, Chief Executive Officer and President, Equinix. “Our longtime relationship with NVIDIA delivers the accelerated computing foundation at the heart of modern AI, while Together AI’s commitment to open ecosystems gives enterprises the flexibility to scale on their terms. Equinix Inference Exchange will enable architectures that are neutral by design, open by default and engineered for exceptional performance.”

“Equinix Inference Exchange turns the world’s leading digital interconnection platform into a global fabric for AI inference,” said Raj Mirpuri, vice president of global AI clouds and infrastructure ecosystem at NVIDIA. “As accelerated compute becomes a strategic asset class, combining NVIDIA’s infrastructure & technology with Together AI’s open-model inference platform and Equinix’s global reach gives enterprises a powerful, distributed foundation to bring intelligence closer to their data, applications and customers – accelerating the next generation of intelligent services.”

“Together AI was built on the conviction that open, accessible AI is what will define the industry moving forward, because enterprises shouldn’t have to choose between model performance and operational flexibility,” said Vipul Ved Prakash, co-founder and CEO, Together AI. “What we are building with Equinix and NVIDIA proves that model choice and performance are not trade-offs. They are the foundation of enterprise AI done right.”

**Where Inference Runs Matters**

The pace of enterprise AI adoption is outrunning the infrastructure required to support it. As enterprise AI shifts from experimentation to production, inference increasingly needs to run closer to the users, data, and applications it serves across clouds, models, providers, and geographies. This transition requires organizations to determine not only how to deploy AI infrastructure, but where it should operate and how it connects to the data, applications, and workloads it relies upon. Managing these distributed inference deployments introduces significant operational complexity at precisely the moment enterprises require greater control and visibility.

“Performance, cost and governance have become strategic considerations as AI workloads grow more distributed across providers, data sources and environments,” said Nick Patience, Vice President & Practice Lead, AI Platforms, The Futurum Group. “Organisations are increasingly focused on where inference runs and how quickly it can be deployed into production. Solutions that simplify inference deployment while preserving flexibility will become increasingly important to achieve business outcomes.”

Equinix brings unmatched scale and ecosystem density to this challenge, operating more than 280 data centers across 77 metros, 230 cloud on-ramps, and over 10,500 interconnected businesses on its neutral exchange. Eight of the top 10 AI model providers and nine of the top 10 AI clouds are deployed within Equinix, underscoring the company’s central position in the AI ecosystem.

**Built for Choice and Flexibility**

Together AI is the latest addition to Equinix’s expansive AI ecosystem, delivering open-model flexibility and choice to enterprises deploying AI at scale. The solution integrates three complementary layers designed to simplify distributed AI inference:

• Equinix provides the infrastructure foundation, including power, advanced cooling, and day-two operations, connected through Equinix Fabric to the clouds, networks, and AI providers that inference depends on.

• NVIDIA anchors the architecture with its Enterprise Reference Architectures and AI infrastructure purpose-built to maximize AI factory throughput and minimize token cost.

• Together AI operates the platform layer, supporting both multitenant deployments for shared efficiency and dedicated single-tenant environments for workloads requiring isolated capacity.

Built on Equinix Fabric, the solution will connect to inference providers across major metros worldwide, significantly reducing time-to-first-token. It also links to an expansive ecosystem of clouds, networks, and AI providers, streamlining deployment complexity.

**Designed for Modern Enterprise Inference**

The solution is engineered to support a broad spectrum of enterprise inference scenarios:

Metro edge inference enables organizations to run inference closer to end users and data sources, delivering lower-latency AI experiences while leveraging Equinix’s security, operational scale, and global reach.

Open model migration provides enterprises transitioning from closed, proprietary models to open-source alternatives with a direct, low-friction path to production. Organizations can control costs and avoid vendor lock-in while running migrations on Together AI’s open-model platform, accessible over the same interconnected fabric used for other providers.

Sovereign AI supports enterprises in regulated industries or specific geographies by enabling workloads to run in locations that satisfy data residency and sovereignty requirements. This offers a streamlined path to scaling AI while maintaining strict control over where data and inference are processed.

Equinix Inference Exchange will be available starting in Q1 2027.

**Additional Resources**

Token Optimization Begins with Choice [Analyst Report]

The Coordination Economy: How Enterprises Really Build AI Value [Blog]

Equinix Inference Exchange [Product Page]

Equinix Inference Exchange Product Release Note [Product Release Note]

Equinix Horizon Event Page [Event Page]

원문 보기