2.9 High-speed network options

NCA-AIIO · AI Infrastructure (40% of the exam) · Official objective: “Identify high speed DC network options and their use cases”

NVLink and NVLink Switch inside a system; InfiniBand and Spectrum-X Ethernet between systems.

Key points

  1. NVLink is NVIDIA's direct GPU-to-GPU interconnect. It scales multi-GPU input and output within a server. NVIDIA lists 1.5x higher bandwidth per GPU at 900 GBps with fourth-generation NVLink on DGX H100. Newer generations are faster, so always check which generation a question means.

    What NVIDIA says (2)

    “is a direct GPU-to-GPU interconnect that scales multi-GPU input/output (IO) in the server.”

    — NVIDIA Fabric Manager User Guide

    “1.5X higher bandwidth per GPU @ 900 GBps with fourth generation of NVIDIA NVLink.”

    — DGX SuperPOD H100 Reference Architecture: Components

  2. Point-to-point links only connect pairs. A switch lets every GPU reach every other GPU. NVIDIA says NVLink Switch chips connect multiple NVLinks to provide all-to-all GPU communication at full NVLink speed. NVLink Switches also include SHARP engines for in-network reductions.

    What NVIDIA says (3)

    “which connects multiple NVLinks to provide all-to-all GPU communication at the total NVLink speed.”

    — NVIDIA Fabric Manager User Guide

    “The NVIDIA NVLink Switch chips connect multiple NVLinks to provide all-to-all GPU communication at full NVLink speed across the entire rack.”

    — NVIDIA NVLink and NVLink Switch

    “To enable high-speed, collective operations, each NVLink Switch has engines for NVIDIA Scalable Hierarchical Aggregation and Reduction Protocol (SHARP)™ for in-network reductions and multicast acceleration.”

    — NVIDIA NVLink and NVLink Switch

  3. Multi-tenant means many customers share the same infrastructure. 'Noise' is one tenant's traffic slowing another. NVIDIA says Spectrum-X Ethernet is designed for cloud providers and large enterprises running multi-tenant AI at hyperscale. It uses telemetry-based congestion control for noise isolation, and SuperNICs provide RoCE (RDMA over Converged Ethernet) between GPU servers.

    What NVIDIA says (4)

    “NVIDIA Spectrum-X is an AI-optimized Ethernet networking platform that combines NVIDIA Spectrum switches with the BlueField-3 SuperNIC”

    — NVIDIA Network Operator: Spectrum-X Ethernet Networking Platform

    “NVIDIA Spectrum-X Ethernet is designed for cloud service providers and large enterprises running multi-tenant AI workloads at hyperscale.”

    — NVIDIA Spectrum-X Ethernet

    “Spectrum-X Ethernet uses advanced telemetry-based congestion control to provide noise isolation between tenants.”

    — NVIDIA Spectrum-X Ethernet

    “the BlueField-3 SuperNIC provides best-in-class remote direct-memory access over converged Ethernet (RoCE) network connectivity between GPU servers”

    — NVIDIA BlueField-3 Networking Platform User Guide: Introduction

  4. SHARP (Scalable Hierarchical Aggregation and Reduction Protocol) lets switches do part of a collective operation, so less data crosses the network. Adaptive routing sends traffic around busy links. NVIDIA highlights SHARP, quality-of-service features including congestion control and adaptive routing, and self-healing networking for Quantum InfiniBand.

    What NVIDIA says (4)

    “technology improves the performance of MPI and Machine Learning collective operation, by offloading collective operations from CPUs and GPUs to the network”

    — NVIDIA SHARP Documentation: Introduction

    “This innovative approach decreases the amount of data traversing the network as aggregation nodes are reached”

    — NVIDIA SHARP Documentation: Introduction

    “NVIDIA Quantum InfiniBand is the only high-performance interconnect solution with proven quality-of-service capabilities, including advanced congestion control and adaptive routing, resulting in unmatched network efficiency.”

    — NVIDIA InfiniBand Networking

    “NVIDIA Quantum InfiniBand with self-healing network capabilities overcomes link failures”

    — NVIDIA InfiniBand Networking

Key terms

Try it

Sample question

What is NVLink, and how fast is it on a DGX H100?

Show the answer

Answer: A direct GPU-to-GPU interconnect inside a server. Fourth-generation NVLink on H100 gives 900 GB/s of bandwidth per GPU.

NVLink is NVIDIA's direct GPU-to-GPU interconnect. It scales multi-GPU input and output within a server. NVIDIA lists 1.5x higher bandwidth per GPU at 900 GBps with fourth-generation NVLink on DGX H100. Newer generations are faster, so always check which generation a question means.

What NVIDIA says (2)

“is a direct GPU-to-GPU interconnect that scales multi-GPU input/output (IO) in the server.”

— NVIDIA Fabric Manager User Guide

“1.5X higher bandwidth per GPU @ 900 GBps with fourth generation of NVIDIA NVLink.”

— DGX SuperPOD H100 Reference Architecture: Components

Practice 2.9 (4 questions) Full AI Infrastructure guide

← 2.8 Data center networking protocols · 2.10 DPUs in the data center →