Equinix Accelerates AI Inference for Enterprises with NVIDIA and Together AI

Equinix, Inc. today announced a significant expansion of its longtime collaboration with NVIDIA to deliver Equinix® Inference Exchange, a distributed AI inference program for global enterprises, alongside a new collaboration with Together AI.

As AI scales across models, providers and geographies, where inference runs is a strategic imperative that determines performance, cost and governance. Equinix Inference Exchange will give enterprises a faster path from AI experimentation to production, with secure, low-latency connectivity to the data, users and ecosystem they depend on.

This collaboration brings together NVIDIA’s validated Enterprise Reference Architectures with Together AI’s inference platform, supporting more than 200 open-source models. Delivered through Equinix’s global data centers, it will provide connectivity to clouds, networks and AI providers through Equinix Fabric®.

The solution will be announced today at Equinix Horizon, the company’s inaugural customer and partner event, alongside Equinix® Fabric One™, which will make it easier for enterprises to connect across globally distributed AI environments.

Where Inference Runs Matters

The pace of enterprise AI adoption is outrunning the infrastructure needed to support it. As enterprise AI moves from experimentation to production, inference increasingly needs to run closer to the users, data and applications it serves across clouds, models, providers and geographies. This shift requires enterprises to determine not only how to deploy AI infrastructure, but where it should run and how it connects to the data, applications and workloads it depends on.

Managing these distributed inference deployments introduces significant operational complexity at precisely the moment enterprises need greater control and visibility.

Equinix brings unmatched scale and ecosystem density to this challenge, with more than 280 data centers across 77 metros, 230 cloud on-ramps and over 10,500 businesses interconnected on its neutral exchange. Eight of the top 10 AI model providers and nine of the top 10 AI clouds are deployed with Equinix, underscoring the company’s position at the center of the AI ecosystem.

Built for Choice and Flexibility

Together AI is the latest addition to Equinix’s expansive AI ecosystem, bringing open-model flexibility and choice to enterprises deploying AI at scale. The solution combines three complementary layers designed to simplify distributed AI inference:

  • Equinix provides the infrastructure foundation, including power, advanced cooling and day-two operations, connected through Equinix Fabric to the clouds, networks and AI providers that inference depends on.
  • NVIDIA anchors the build with its Enterprise Reference Architectures and AI infrastructure purpose-built to maximize AI factory throughput and minimize token cost.  
  • Together AI runs the platform on top, supporting both multitenant deployments for shared efficiency and dedicated single-tenant environments for workloads that require dedicated capacity.

Built on Equinix Fabric, the solution will connect to inference providers across major metros worldwide, cutting time-to-first-token. It also will connect to an expansive ecosystem of clouds, networks and AI providers, reducing deployment complexity.

Designed for Modern Enterprise Inference

The solution aims to support a broad range of enterprise inference scenarios, including:

  • Metro edge inference: For organizations that need inference running closer to users and data, enabling lower-latency AI experiences while leveraging the security, operational scale and global reach of Equinix.
  • Open model migration: For enterprises moving workloads from closed, proprietary models to open-source alternatives to control cost and avoid lock-in, the solution will provide a direct, low-friction path to run that migration in production, with Together AI’s open-model platform reachable over the same interconnected fabric enterprises already use to reach their other providers.
  • Sovereign AI: For enterprises operating in regulated industries or specific geographies, the solution will enable AI workloads to run in locations that support data residency and sovereignty requirements, providing a simpler path to deploying AI at scale while maintaining control over where data and inference are processed.

Equinix Inference Exchange will be available starting in Q1 2027.

Leave a Reply

Discover more from The IT Nerd

Subscribe now to keep reading and get access to the full archive.

Continue reading