

A fictional organization is building an AI training and distributed-inference environment with 512 GPU servers across 16 racks. Each server has one 800G-class AI-network attachment represented here as two 400G ports for simple cabling arithmetic, plus separate service and OOB connectivity. Storage is logically isolated and may be physically separated as growth requires.
Use a two-tier leaf-spine design. Each rack receives redundant leaf connectivity or independent rails according to the selected Spectrum-X reference architecture. Spines provide equal-cost paths between every rack. If the selected generation/radix cannot meet growth targets in two tiers, evaluate Spectrum-X Multiplane before adding a third tier.

Figure 30. End-to-end fictional AI factory
For each rack, calculate total server injection bandwidth and match leaf uplink bandwidth to the target oversubscription ratio. For training-heavy workloads, begin with a non-blocking or close-to-non-blocking target, then validate with application models. Reserve headroom so loss of one spine or one rail does not immediately create persistent queues.
Map each GPU server’s SuperNIC ports to independent leaf/plane paths. Confirm PCIe and GPU locality so the NIC serving a GPU is not unnecessarily crossing CPU sockets or constrained PCIe paths. Use the NVIDIA-validated NIC firmware and Spectrum-X profile for the selected reference architecture.
Run an L3 underlay with eBGP leaf-to-spine adjacencies and ECMP. Use loopbacks for stable identities and BFD where the failure-detection requirement justifies it. Keep the AI fabric routing table intentionally simple: host/rail prefixes and infrastructure reachability, not enterprise service routes.
Use the validated Spectrum-X RoCE profile. ECN and endpoint congestion control should prevent persistent queues; PFC, if the validated mode uses it, is limited to the required priority. Enable supported adaptive routing on the eligible links and observe path-balance telemetry. Do not hand-tune thresholds before establishing a repeatable workload baseline.
Model link, leaf, spine, and rack failures. Confirm both logical reachability and remaining bandwidth. Define a maintenance drain procedure so a spine or leaf can be removed from service without surprising the application team.
All configuration originates from source control. CI validates addressing, BGP peers, MTU, QoS profiles, and topology. DSX Air validates control-plane and automation behaviour. NetQ and external telemetry monitor path imbalance, queues, RoCE, physical health, and routing. Application dashboards track NCCL bandwidth and GPU utilization so network events can be correlated to business-relevant performance.
When rack count or GPU count approaches the two-tier radix limit, evaluate larger Spectrum generations, additional planes, or pod expansion. The decision should be based on effective bandwidth, failure domains, optics/power, and operational complexity rather than a desire to keep the topology visually simple.
Source note: See references [1], [3], [16], [27], [31].
| Dimension | Traditional enterprise network | Spectrum / AI-scale fabric |
| Traffic pattern | User/application, north-south heavy | Machine-to-machine, east-west heavy |
| Endpoint speed | 1/10/25G common | 100/200/400/800G-class server/uplink links |
| Topology | Access/distribution/core | Clos leaf-spine, multi-rail/multiplane at scale |
| Oversubscription | Often high and acceptable | Training may require low oversubscription |
| Latency sensitivity | Application dependent | Collective phases can amplify tail latency |
| Congestion sensitivity | TCP usually masks moderate congestion | Congestion can directly idle expensive GPUs |
| Routing | OSPF/BGP/static mix; L2 common at access | L3 BGP/ECMP underlay common |
| Overlay | Campus segmentation or DC virtualization | EVPN/VXLAN when multi-tenancy needs it; optional for AI back-end |
| Automation | Often partial | Essential at large scale |
| Observability | Device/interface monitoring | Queue/path/RoCE/application correlation |
| Failure handling | Redundant devices and protocols | Parallel paths plus capacity headroom |
| RDMA | Rare | Common for high-performance compute/storage |
| Term | Beginner-friendly definition |
| ACL | Access Control List; rules that permit, deny, or classify traffic. |
| ASIC | Application-Specific Integrated Circuit; the switch silicon that forwards packets at line rate. |
| BFD | Bidirectional Forwarding Detection; a fast failure-detection protocol commonly paired with routing. |
| BGP | Border Gateway Protocol; scalable routing protocol widely used in data-centre leaf-spine fabrics. |
| Bisection bandwidth | Aggregate bandwidth available between two halves of a network; useful for assessing distributed-workload capacity. |
| Clos | Multi-stage network topology that provides many parallel paths. Leaf-spine is a common two-stage Clos form. |
| CNP | Congestion Notification Packet used in RoCE congestion-control feedback. |
| Cumulus Linux | NVIDIA Debian-based network operating system for Spectrum switches. |
| DCQCN | Data Center Quantized Congestion Notification; a common RoCEv2 congestion-control algorithm using ECN feedback. |
| DPU | Data Processing Unit; programmable infrastructure processor for networking, storage, and security offload. |
| DSX Air | NVIDIA cloud-hosted data-centre simulation/digital-twin platform. |
| ECMP | Equal-Cost Multi-Path; forwarding across multiple routes with equal routing cost. |
| ECN | Explicit Congestion Notification; marks packets to signal congestion without dropping them. |
| EVPN | Ethernet VPN; BGP-based control plane often used with VXLAN to distribute endpoint/tenant reachability. |
| GPUDirect RDMA | Technology enabling RDMA-capable adapters to transfer data directly to/from GPU memory. |
| Leaf | Switch tier connected to servers or endpoints; each leaf connects upward to all spines in a classic leaf-spine fabric. |
| MLAG | Multi-Chassis Link Aggregation; two switches present an active-active LAG to attached devices. |
| NCCL | NVIDIA Collective Communications Library; GPU collective communications library used by distributed AI applications. |
| NetQ | NVIDIA network operations and telemetry platform for Cumulus/Spectrum environments. |
| NVUE | NVIDIA User Experience; structured configuration and operational interface used by Cumulus Linux. |
| PFC | Priority Flow Control; Ethernet mechanism that pauses selected traffic priorities hop by hop. |
| RDMA | Remote Direct Memory Access; direct memory-to-memory data transfer with reduced CPU involvement. |
| RoCE | RDMA over Converged Ethernet. |
| RoCEv2 | Routable RoCE transport over UDP/IP. |
| Spine | Switch tier that interconnects all leaf switches in a leaf-spine fabric. |
| Spectrum | NVIDIA family of high-performance Ethernet switching ASICs and switch systems. |
| Spectrum-X | NVIDIA AI-optimized Ethernet platform combining Spectrum switches, SuperNICs, software, telemetry, and validated tuning. |
| Spectrum-XGS | Spectrum-X scale-across technology for connecting distributed data centres into a larger AI factory. |
| SuperNIC | NVIDIA term for a network accelerator optimized for network-intensive AI workloads. |
| ToR | Top of Rack; a switch physically located in or associated with a server rack, often acting as a leaf. |
| VNI | VXLAN Network Identifier; identifies a logical VXLAN segment. |
| VRF | Virtual Routing and Forwarding instance; creates separate routing tables for isolation. |
| VTEP | VXLAN Tunnel Endpoint; device that encapsulates/decapsulates VXLAN traffic. |
| VXLAN | Virtual Extensible LAN; UDP-based overlay encapsulation used to carry tenant networks over an IP underlay. |
Primary technical references are NVIDIA sources because product positioning, supported combinations, and release qualifications change quickly. Standards references should be added when turning individual sections into deeply cited blog posts.
[1] NVIDIA, “NVIDIA Spectrum-X Ethernet Networking Platform,” current product overview, accessed September 2026. https://www.nvidia.com/en-us/networking/spectrumx/
[2] NVIDIA, “Ethernet Switching for AI and the Cloud,” Spectrum Ethernet switch portfolio, accessed September 2026. https://www.nvidia.com/en-us/networking/ethernet-switching/
[3] NVIDIA Docs, “NVIDIA Spectrum-X Ethernet Networking Platform” (Kubernetes/Network Operator documentation), release 26.7 context. https://docs.nvidia.com/networking/display/kubernetes2670/spectrum-x/spectrum-x.html
[4] NVIDIA, “Networking Solutions for the Era of AI,” Ethernet, InfiniBand, and BlueField portfolio overview. https://www.nvidia.com/en-us/networking/
[5] NVIDIA Docs, “RDMA over Converged Ethernet – RoCE,” Cumulus Linux documentation. https://docs.nvidia.com/networking-ethernet-software/cumulus-linux-515/Layer-1-and-Switch-Ports/Quality-of-Service/RDMA-over-Converged-Ethernet-RoCE/
[6] NVIDIA Docs, “Quality of Service,” Cumulus Linux 5.18, including PFC considerations. https://docs.nvidia.com/networking-ethernet-software/cumulus-linux/Layer-1-and-Switch-Ports/Quality-of-Service/
[7] NVIDIA Newsroom, “NVIDIA Completes Acquisition of Mellanox,” 27 April 2020. https://nvidianews.nvidia.com/news/nvidia-completes-acquisition-of-mellanox-creating-major-force-driving-next-gen-data-centers
[8] NVIDIA Newsroom, “NVIDIA Announces Spectrum High-Performance Data Center Networking Infrastructure Platform,” 22 March 2022. https://nvidianews.nvidia.com/news/nvidia-announces-spectrum-high-performance-data-center-networking-infrastructure-platform
[9] NVIDIA Blog, “Built for Vera Rubin, NVIDIA Spectrum-6 Arrives in Gigascale AI Factories,” 21 July 2026. https://blogs.nvidia.com/blog/nvidia-spectrum-six-arrives-in-gigascale-ai-factories/
[10] NVIDIA Docs, “NVIDIA Networking Documentation,” Cumulus Linux, NetQ, and DSX Air documentation hub. https://docs.nvidia.com/networking-ethernet-software/
[11] NVIDIA Docs, “Cumulus Linux Release Versioning and Support Policy,” accessed September 2026. https://docs.nvidia.com/networking-ethernet-software/knowledge-base/Support/Support-Offerings/Cumulus-Linux-Release-Versioning-and-Support-Policy/
[12] NVIDIA Docs, “Cumulus Linux 5.15 User Guide” and current 5.x documentation. https://docs.nvidia.com/networking-ethernet-software/cumulus-linux-515/
[13] NVIDIA, “BlueField Networking Platform,” BlueField-3 and BlueField-4 overview. https://www.nvidia.com/en-us/networking/products/data-processing-unit/
[14] NVIDIA Docs, “NVIDIA DSX Air User Guide.” https://docs.nvidia.com/networking-ethernet-software/nvidia-air/
[15] NVIDIA Newsroom, “NVIDIA Introduces Spectrum-XGS Ethernet,” 22 August 2025. https://nvidianews.nvidia.com/news/nvidia-introduces-spectrum-xgs-ethernet-to-connect-distributed-data-centers-into-giga-scale-ai-super-factories
[16] NVIDIA, “Spectrum-X Validated Solution Stack,” August 2026 v2.1.6 and prior validated combinations. https://networking-docs.nvidia.com/software/spectrumx-solution-stack
[17] NVIDIA, “NVIDIA InfiniBand Adapters” and current InfiniBand platform material. https://www.nvidia.com/en-us/networking/infiniband-adapters/
[18] NVIDIA Technical Blog, “How to Connect Distributed Data Centers Into Large AI Factories with Scale-Across Networking,” September 2025. https://developer.nvidia.com/blog/how-to-connect-distributed-data-centers-into-large-ai-factories-with-scale-across-networking/
[19] NVIDIA NCCL documentation, current release documentation. https://docs.nvidia.com/deeplearning/nccl/
[20] NVIDIA Docs, Spectrum-X NIC Configuration, Network Operator 26.7. https://docs.nvidia.com/networking/display/kubernetes2670/spectrum-x/spectrum-x-configuration.html
[21] NVIDIA Docs, Adaptive Routing, NVUE 5.x. https://docs.nvidia.com/networking-ethernet-software/nvue-reference/Set-and-Unset-Commands/Adaptive-Routing/
[22] NVIDIA Docs, Cumulus Linux EVPN/VXLAN active-active mode. https://docs.nvidia.com/networking-ethernet-software/cumulus-linux/Network-Virtualization/VXLAN-Active-Active-Mode/
[23] NVIDIA Docs, Cumulus Linux EVPN inter-subnet routing. https://docs.nvidia.com/networking-ethernet-software/cumulus-linux-515/Network-Virtualization/Ethernet-Virtual-Private-Network-EVPN/Inter-subnet-Routing/
[24] NVIDIA Docs, NVUE CLI, Cumulus Linux 5.x. https://docs.nvidia.com/networking-ethernet-software/cumulus-linux-515/System-Configuration/NVIDIA-User-Experience-NVUE/NVUE-CLI/
[25] NVIDIA, “Onyx for Next-Generation Data Centers,” product page. https://www.nvidia.com/en-au/networking/ethernet-switching/onyx/
[26] NVIDIA Spectrum-6 SN6000 Ethernet Switch Systems Hardware User Manual, 2026. https://docs.nvidia.com/nvidia-spectrum-6-sn6000-ethernet-switch-systems-hardware-user-manual.pdf
[27] NVIDIA Docs, DSX Air Custom Topology and Quick Start. https://docs.nvidia.com/networking-ethernet-software/nvidia-air/Custom-Topology/
[28] NVIDIA Docs, NVIDIA NetQ 5.3 User Guide. https://docs.nvidia.com/networking-ethernet-software/cumulus-netq-53/
[29] NVIDIA Docs, Cumulus Linux configuration guidance for Ethernet Storage Fabrics, including MLAG/active-active design. https://docs.nvidia.com/networking-ethernet-software/guides/esf-generic-config-guide/
[30] NVIDIA Spectrum switch hardware manuals and Ethernet switching product tables, current generations. https://www.nvidia.com/en-us/networking/ethernet-switching/
[31] NVIDIA Docs, Spectrum-X Launch Kit / Network Operator profiles, current release documentation. https://docs.nvidia.com/networking/display/kubernetes2670/k8s-launch-kit/profiles/spectrum-x.html
[32] NVIDIA Docs, Cumulus Linux authentication, authorization, user accounts, and system configuration documentation. https://docs.nvidia.com/networking-ethernet-software/cumulus-linux/System-Configuration/Authentication-Authorization-and-Accounting/User-Accounts/
[33] NVIDIA Docs, NetQ release/version support policy and current NetQ 5.3 documentation. https://docs.nvidia.com/networking-ethernet-software/knowledge-base/Support/Support-Offerings/Cumulus-NetQ-Release-Versioning-and-Support-Policy/