Home / _ / NVIDIA DGX H100 AI Supercomputer

NVIDIA DGX H100 AI Supercomputer

 

The NVIDIA DGX H100 is an enterprise-class AI supercomputer designed for large-scale artificial intelligence, generative AI, deep learning, high-performance computing (HPC), data analytics, and large language model workloads.

Powered by 8 NVIDIA H100 Tensor Core GPUs, DGX H100 provides 640GB of total GPU memory, connected through NVIDIA’s fourth-generation NVLink and NVSwitch architecture. The system delivers up to 32 PFLOPS of FP8 AI performance and provides high-bandwidth networking through NVIDIA ConnectX-7 adapters.

The system combines GPU acceleration, dual Intel Xeon processors, 2TB of system memory, high-speed NVMe storage, ConnectX-7 networking, NVSwitch fabric, and NVIDIA’s optimized DGX software stack into a fully integrated AI infrastructure platform.


Category:

Description

The NVIDIA DGX H100 is the fourth-generation NVIDIA DGX system and was specifically engineered to provide a complete infrastructure platform for enterprise-scale AI.

Rather than requiring organizations to integrate GPUs, networking, storage, and software independently, DGX H100 combines these components into a validated system optimized for large distributed AI workloads.

Eight NVIDIA H100 Tensor Core GPUs

At the heart of DGX H100 are 8 NVIDIA H100 Tensor Core GPUs based on the NVIDIA Hopper architecture.

Together, the eight GPUs provide:

  • 640GB total HBM3 GPU memory
  • 80GB HBM3 per GPU
  • Up to 3.35TB/s memory bandwidth per GPU
  • 8,192-bit aggregate GPU memory interface
  • Fourth-generation Tensor Cores
  • Fourth-generation NVLink connectivity

The H100 GPU is optimized for transformer-based AI, generative AI, scientific computing, deep learning, recommendation systems, and other computationally intensive workloads. NVIDIA specifies up to 34 TFLOPS FP64, 67 TFLOPS FP32, and 3,958 TFLOPS FP8 Tensor Core performance with sparsity for the H100 SXM configuration.

NVIDIA NVLink and NVSwitch

The eight GPUs are interconnected using NVIDIA’s high-speed GPU fabric.

DGX H100 incorporates:

  • 4 × fourth-generation NVIDIA NVSwitch
  • 18 NVLink connections per GPU
  • Up to 900GB/s bidirectional GPU-to-GPU bandwidth per GPU
  • Up to 7.2TB/s aggregate bidirectional GPU-to-GPU bandwidth

This architecture allows the eight H100 GPUs to operate as a tightly interconnected accelerated computing platform rather than as independent graphics processors.

High-Performance Networking

DGX H100 integrates NVIDIA ConnectX-7 networking for high-speed cluster communication.

The system provides:

  • 8 × single-port ConnectX-7 400Gb/s InfiniBand adapters
  • 2 × dual-port ConnectX-7 Ethernet adapters
  • Up to 400Gb/s InfiniBand
  • Up to 400GbE Ethernet
  • Support for 200GbE, 100GbE, 50GbE, 40GbE, 25GbE and 10GbE depending on interface/configuration

The networking architecture provides up to approximately 1TB/s of peak bidirectional network bandwidth across the system.

This makes DGX H100 suitable for multi-node AI clusters, DGX SuperPOD environments, distributed training, large-scale inference, and high-performance storage.

Dual Intel Xeon Platinum Processors

DGX H100 uses two Intel Xeon Platinum 8480C processors.

Together they provide:

  • 112 CPU cores total
  • 56 cores per processor
  • Base frequency: 2.0GHz
  • All-core turbo: up to 2.9GHz
  • Maximum turbo: up to 3.8GHz

The CPU subsystem handles system management, data preparation, orchestration, storage, networking, and host-side application workloads.

2TB System Memory

The system includes 2TB of system DDR5 memory using 32 DIMMs.

This large CPU memory capacity supports:

  • Large datasets
  • Data preprocessing
  • AI training pipelines
  • Multi-GPU workloads
  • HPC applications
  • High-performance databases
  • Large-scale analytics

High-Speed NVMe Storage

DGX H100 includes approximately 30TB of internal NVMe storage in the standard configuration.

The storage architecture consists of:

  • 2 × 1.92TB NVMe M.2 SSDs configured as RAID 1 for the operating system
  • 8 × 3.84TB NVMe U.2 self-encrypting drives configured as RAID 0 for data cache

This provides fast local storage for operating-system files, datasets, model checkpoints, temporary data, and AI workloads.

BlueField-3 and Infrastructure Acceleration

DGX H100 incorporates NVIDIA networking and infrastructure acceleration technologies designed to offload networking, storage, and security services.

NVIDIA’s original DGX H100 announcement also identifies two NVIDIA BlueField-3 DPUs within the system architecture for infrastructure offload and isolation.

Enterprise Software Stack

DGX H100 is designed to operate with NVIDIA’s optimized DGX software environment.

The platform supports:

  • NVIDIA DGX OS
  • CUDA
  • NVIDIA GPU drivers
  • NVIDIA Base Command
  • NVIDIA AI Enterprise
  • NVIDIA Magnum IO
  • NVIDIA networking software
  • GPU Direct Storage
  • AI and HPC frameworks

Current NVIDIA documentation lists DGX H100 as supported by DGX OS 8, while DGX OS 7 also supports the system.


3. Complete Technical Specification

Specification Details
Product Name NVIDIA DGX H100
Product Type Enterprise AI Supercomputer / AI Server
GPU Architecture NVIDIA Hopper
GPU 8 × NVIDIA H100 Tensor Core GPU
GPU Form Factor SXM
GPU Memory 640GB total
Memory per GPU 80GB HBM3
GPU Memory Bandwidth Up to 3.35TB/s per GPU
GPU-to-GPU Interconnect 4th-generation NVIDIA NVLink
NVLink per GPU 18 connections
GPU-to-GPU Bandwidth Up to 900GB/s bidirectional per GPU
NVSwitch 4 × NVIDIA NVSwitch
Aggregate GPU Fabric Bandwidth Up to 7.2TB/s bidirectional
FP8 AI Performance Up to 32 PFLOPS with sparsity
CPU 2 × Intel Xeon Platinum 8480C
CPU Cores 112 total / 56 per CPU
CPU Base Frequency 2.0GHz
CPU All-Core Turbo Up to 2.9GHz
CPU Maximum Turbo Up to 3.8GHz
System Memory 2TB
Memory Configuration 32 × DIMMs
Cluster Networking 8 × ConnectX-7 single-port 400Gb/s adapters
Cluster Network Interface OSFP
InfiniBand Up to 400Gb/s
Ethernet Networking Up to 400GbE
Ethernet Speeds 400/200/100/50/40/25/10GbE
Storage Networking 2 × ConnectX-7 dual-port Ethernet adapters
OS Storage 2 × 1.92TB NVMe M.2
OS RAID RAID 1
Data Cache Storage 8 × 3.84TB NVMe U.2 SED
Data RAID RAID 0
Total Internal NVMe Capacity Approximately 30TB
BMC Integrated
BMC Network 1GbE RJ45
Management Protocols Redfish, IPMI, SNMP, KVM, Web UI
Power Supplies 6 × 3.3kW
Maximum System Power Approximately 10.2kW
Operating Temperature 5°C–30°C
System Architecture x86-64
Operating System NVIDIA DGX OS
Primary Software CUDA, NVIDIA AI Enterprise, Base Command, Magnum IO
Primary Workloads AI, Generative AI, LLMs, HPC, analytics
System Class Enterprise data-center infrastructure

The current NVIDIA user guide confirms the 8-H100/640GB configuration, dual Xeon 8480C CPUs, 2TB system memory, 30TB-class NVMe storage, ConnectX-7 networking, four NVSwitches, and six 3.3kW power supplies.


4. Applications

Artificial Intelligence & Machine Learning

  • Large language model training
  • Generative AI
  • Transformer models
  • Deep learning
  • Recommendation systems
  • Computer vision
  • Speech and language AI
  • Multimodal AI
  • Large-scale inference

High-Performance Computing

  • Scientific simulations
  • Climate modeling
  • Molecular dynamics
  • Computational chemistry
  • Engineering simulation
  • Physics research
  • Financial modeling
  • Genomics and life sciences

Enterprise AI

  • Enterprise AI platforms
  • AI-powered analytics
  • Intelligent automation
  • Natural language processing
  • Customer intelligence
  • Recommendation engines
  • Predictive analytics

AI Infrastructure

  • NVIDIA DGX SuperPOD
  • AI training clusters
  • Multi-node GPU infrastructure
  • Private AI clouds
  • Enterprise AI factories
  • Research supercomputing

5. Compatibility

The DGX H100 is a complete integrated data-center system, rather than a conventional server intended for component-level GPU upgrades.

It is designed for deployment with:

  • NVIDIA DGX SuperPOD infrastructure
  • NVIDIA DGX BasePOD architectures
  • InfiniBand networks
  • 400GbE networks
  • NVIDIA Quantum networking
  • NVIDIA ConnectX-7 networking
  • High-performance NVMe storage
  • NVIDIA AI Enterprise
  • NVIDIA Base Command
  • CUDA-based applications
  • Linux AI/HPC environments
  • Enterprise AI clusters

For cluster deployment, the system is designed to connect through its high-speed ConnectX-7 networking to other DGX systems and compatible NVIDIA networking infrastructure.

Important: DGX H100 should not be confused with a standard 8-GPU HGX H100 server. DGX H100 is an NVIDIA-integrated system with validated GPU, NVSwitch, networking, storage, CPU, software, and management components.


6. Key Benefits

  • 8 × NVIDIA H100 Tensor Core GPUs
  • 640GB total HBM3 GPU memory
  • Up to 32 PFLOPS FP8 AI performance
  • High-bandwidth 4th-generation NVLink
  • 4 × NVIDIA NVSwitch
  • Up to 7.2TB/s aggregate GPU fabric bandwidth
  • 112 Intel Xeon CPU cores
  • 2TB system memory
  • Approximately 30TB NVMe storage
  • 400Gb/s ConnectX-7 networking
  • InfiniBand and Ethernet connectivity
  • Integrated BMC and enterprise management
  • Designed for NVIDIA DGX SuperPOD environments
  • Optimized for generative AI and large language models
  • Enterprise-grade AI infrastructure
  • Integrated NVIDIA software ecosystem
  • High-speed GPU, network, and storage data paths
  • Suitable for demanding AI, HPC and analytics workloads

7. SEO Title

NVIDIA DGX H100 AI Supercomputer – 8× H100 640GB, 2TB RAM, 400Gb/s Networking