Accelerating Ai Networking With Nvidia Spectrum X And Aviz Ones - Resource | Aviz Networks | Aviz Networks

ONES + NVIDIA Spectrum-X Ethernet | Intelligent AI Fabric Orchestration

A unified platform to design, deploy, and operate AI-ready Ethernet fabrics

February 2, 2025

Overview

The ROI of AI and HPC clusters depends on the high-performance network fabric that supports them. Training large-scale AI models and deploying multi-tenant AI inference clusters requires ultra-fast GPU-to-GPU and GPU-to-storage communication, ultra-low latency, lossless data transfer, and seamless scalability. NVIDIA Spectrum-X Ethernet provides this foundation with Spectrum-4 Ethernet switches, NVIDIA SuperNICs, and NVIDIA GPU Platforms.

Aviz ONES (Open Networking Enterprise Suite) complements Spectrum-X Ethernet by providing intelligent multi-tenant orchestration, simplifying day-to-day operations, and enabling comprehensive monitoring across the fabric. ONES also supports customers to expand their AI factory footprint over time without impacting existing tenants. Together, ONES and Spectrum-X Ethernet allow enterprises to achieve predictable AI performance with greater simplicity and confidence.

Spectrum-X Ethernet Fabric Architecture

ONES is fully aligned with the NVIDIA Spectrum-X Ethernet Reference Architecture (RA). At present, ONES supports RA 1.3, which is built on:

01

NVIDIA GPU Platforms

The industry's leading AI compute platform.

02

Spectrum-4 Ethernet Switches

High-density Ethernet switches optimized for AI networking.

03

BlueField-3 SuperNICs

Hardware acceleration for RoCEv2 and advanced telemetry.

As NVIDIA evolves the Spectrum-X Ethernet RA, ONES will continue to align with future RA versions and their components — ensuring that enterprises can confidently adopt the latest GPU, switch, and NIC innovations while maintaining consistent orchestration and operational workflows.

Figure 1: NVIDIA Spectrum-X Ethernet Rail Optimized Architecture

This combination delivers high-performance Ethernet networking, providing a lossless, ultra-low latency GPU fabric that is purpose-built to accelerate AI workloads, maximize GPU utilization, and ensure predictable performance at scale.

Orchestrating AI Fabrics with ONES

Traditional data center networks were not built for AI, often struggling with multi-tenancy, lossless transport, and lifecycle operations. A modern AI Factory is rail-optimized, combining an East-West GPU fabric for high-bandwidth compute, a North-South fabric for CPU, storage, and external user access.

The Scalable Unit (SU) is the modular building block combining up to 32 GPU Nodes and the required leaf switches for the North-South and East-West fabrics. The maximum size of a single AI factory deployment with a 2-tier leaf-spine network is 32 SUs with a total of 1,024 GPU Nodes and 8,192 GPUs. ONES orchestrates across all hosts and SUs consistently, enabling smooth deployment and seamless growth from hundreds to thousands of GPUs.

Figure 2: Topology for ONES Spectrum-X Ethernet Orchestration and Deployment

Addressing Network Challenges

Conventional network tools fall short in meeting the demands of AI fabrics. ONES addresses the gap with:

ONES: Lifecycle Orchestration for Spectrum-X Ethernet

Day-0: Orchestration & Design

01

Simplified Inputs

Define SU count, subnets, and storage type. Provide customers with the option to design for future growth, minimizing operational impact when expanding the AI factory footprint.

02

Fabric Intent

Define whether the AI fabric is built only for GPU-to-GPU traffic (East-West) or for GPU traffic plus external ingress/egress (East-West + North-South) at Day-0.

03

Intent-based Design

ONES auto-generates validated configs for underlay, QoS, and NICs.

04

Pre-deployment Validation

Use NVIDIA AIR to simulate topologies before rollout.

05

Zero-Touch Provisioning (ZTP)

Push configs across switches and servers consistently. Provide customers with the option for both greenfield and brownfield deployments (e.g., validate existing configurations).

Figure 3: Fabric Orchestration for East-West (GPU-to-GPU) and North-South (To External and Storage Networks) using ONES Orchestrator

Spectrum-X Ethernet Orchestration

ONES provides a unified interface for designing, viewing, and managing AI factory topologies. The Fabric Designer visualizes your Scalable Units, network layers, and device relationships — making complex AI factory architectures easy to understand and manage.

NVIDIA Air Integration

Validate your AI factory designs end-to-end using NVIDIA Air digital-twin simulation. ONES integrates directly with NVIDIA Air, allowing you to simulate topologies, test configurations, and verify operational readiness before production rollout.

ONES: Day-1 Multi-Tenant Operations

01

Tenant Management

Create isolated tenants using VXLAN/EVPN with policy control.

02

GPU-aware Provisioning

Align network resources with compute.

03

Role-based GUI Orchestration

Simplifies operations without CLI.

Figure 5: GPU Allocations to Tenants using ONES Orchestrator

Day-2: Operations & Integrations

ONES provides seamless integrations with enterprise tools and data platforms through its Data Connector. Connect to ticketing systems like Zendesk and ServiceNow, monitoring platforms like Splunk, and cloud services like Amazon S3 — enabling automated workflows and centralized data management.

Key Benefits

1. Simple Design
Intent-based workflows and AIR validation reduce complexity.

2. Fast Deployment
From weeks to hours with Day-0 orchestration.

3. Secure Multi-Tenancy
GPU-aware VXLAN/EVPN segmentation ensures isolation.

4. Resilient Operations
Full-stack visibility and lifecycle workflows keep fabrics reliable.

The AI-Ready Fabric Advantage

Capability Value
Fabric Automation Intent-driven design validated with NVIDIA AIR
Multi-Tenancy Secure EVPN/VRF segmentation with GPU awareness
Observability Switch, NIC, and GPU telemetry via NetQ APIs
Resilience Seamless scaling and automated RMA workflows

Conclusion

Networking for AI and HPC is about more than throughput — it's about delivering a predictable, lossless, ultra-low latency, and efficiently managed Ethernet fabric for AI Factory performance at scale.

With ONES + Spectrum-X Ethernet, organizations can design, deploy, and operate AI-ready Ethernet fabrics with the speed, simplicity, and resilience required by today's most demanding workloads.

**ONES and Spectrum-X together provide a complete solution for AI fabric orchestration, from initial design through multi-tenant operations to ongoing management and scaling.