Skip to main content
VTechFusion Technologies
Equinix Unveils AI Inference Exchange With Nvidia and Together AI, Launching Q1 2027
InsightsNewsIndustry & AI News
Industry & AI News5 min readSeptember 2, 2026

Equinix Unveils AI Inference Exchange With Nvidia and Together AI, Launching Q1 2027

VT

VTechFusion Team

VTechFusion Technologies

Equinix announced its Inference Exchange on September 2, 2026, combining Nvidia's Enterprise Reference Architectures with Together AI's inference platform and Equinix's global data center footprint, aimed at giving enterprises a faster path from AI experimentation to production deployment. The service is scheduled to launch in the first quarter of 2027.

The Technical Components

  • Nvidia's validated Enterprise Reference Architectures provide the underlying infrastructure design patterns
  • Together AI's inference platform supports more than 200 open-source models, giving enterprises a broad range of model choice within the service
  • Delivery runs through Equinix Fabric, connecting to Equinix's network of more than 280 data centers across 77 metros with over 230 cloud on-ramps

The Strategic Bet: Proximity Matters for Inference

The service is specifically framed around low-latency connectivity to the data, users, and ecosystem an enterprise already depends on — a bet that running inference physically closer to where data and users actually are matters enough to enterprises to build dedicated infrastructure around it, distinct from simply renting more centralized cloud compute. This is consistent with a broader industry pattern of inference workloads (as opposed to training) demanding different infrastructure characteristics, including proximity and latency, not just raw compute scale.

What's Still Unknown

Equinix has not yet disclosed pricing, capacity commitments, or anchor customers ahead of the Q1 2027 launch — meaning organizations interested in the service have real details still to learn before making any commitment. For enterprises currently planning AI infrastructure for 2027, this is worth tracking as a potential option specifically for latency-sensitive inference and agentic workloads, without yet being able to fully evaluate its cost competitiveness against existing cloud inference options.

Filed under:Industry & AI News
All News

Frequently Asked Questions

What is Equinix's Inference Exchange?

A service announced September 2, 2026 combining Nvidia's Enterprise Reference Architectures, Together AI's inference platform (supporting 200+ open models), and Equinix's global data center network, aimed at giving enterprises low-latency AI inference infrastructure. It launches in Q1 2027.

Why does Equinix emphasize proximity and latency for this service?

The service is built around running inference physically closer to an enterprise's own data and users, reflecting a broader industry pattern where inference workloads have different infrastructure needs (proximity, latency) than training workloads, which prioritize raw compute scale.

What details about the Inference Exchange are still unknown?

Pricing, capacity commitments, and anchor customers have not yet been disclosed ahead of the Q1 2027 launch, meaning interested organizations don't yet have enough information to fully evaluate its cost competitiveness.

Media & Press Enquiries

For editorial enquiries, expert commentary, or case study access.

Start Today

Ready to Build Something Great?

Let's turn your idea into a product. Book a free 30-minute discovery call with our team — no commitment, just clarity.