NVIDIA GB200 NVL72

Engineered on NVIDIA’s cutting-edge Blackwell architecture and coupled with the Arm-based Grace CPU, the GB200 NVL72 sets a new standard for AI compute.

It delivers up to 30x faster real-time LLM inference, cuts Total Cost of Ownership (TCO) by 25x, and consumes 25x less energy compared to previous generations.

Pricing details will be announced upon official commercial availability.

Accelerate AI Enterprise Workflows with NVIDIA DGX B300

The NVIDIA DGX B300 serves as a unified AI supercomputing platform designed to empower developers and data science teams.

By dramatically accelerating time-to-insight, it helps organizations unlock the full strategic and economic potential of enterprise AI.

Technical Specifications (DGX B300)

  • Processors: 2x Intel Xeon 6776P (64 Cores, 2.3 GHz)
  • GPU Acceleration: 8x NVIDIA B300 SXM
  • Memory: 2 TB System RAM
  • Storage: 30 TB High-Speed NVMe (Data)
  • High-Speed Interconnect: 8x 400Gb HDR InfiniBand
  • Networking: 4x 400Gb Ethernet
  • Service: Includes 3-Year Enterprise Support Package

 

// solution overview

NVIDIA GB200 NVL72 System

Engineered for massive-scale AI workloads, the NVIDIA GB200 NVL72 effortlessly processes datasets containing trillions of parameters.

Combining the state-of-the-art Blackwell architecture with the high-performance Arm Grace CPU, this advanced superchip delivers unprecedented computing power while slashing both operational costs (TCO) and energy consumption by up to 25x relative to prior-generation architectures.

performance
Configuration flexibility
TCO
premium support

Substantial performance gain with GB200 NVL72

// Key features

LLM Inference

30X

vs H100 Tensor Core GPU*
 

LLM Training

4X

vs H100*
 

Energy Efficiency

25X

vs H100*
 

Data Processing

18X

vs CPU*
 
 

NVIDIA: LLM inference and energy efficiency: Time to First Token (TTFT) = 50 ms real-time, Full-Time Latency (FTL) = 5 s, with 32,768 input and 1,024 output tokens.

Benchmark compares NVIDIA HGX™ H100 scaled over InfiniBand (IB) against the GB200 NVL72; 1.8T MoE model training evaluates 4,096x HGX H100 scaled via IB versus 456x GB200 NVL72 scaled via IB within a 32,768-cluster size.

Database join and aggregation performance with Snappy/Deflate compression is derived from the TPC-H Q4 query, featuring custom query implementations across x86, a single H100 GPU, and a single GPU from the GB200 NVL72 vs.

Intel Xeon 8480+.

Projected performance subject to change.

NVIDIA GB200 NVL72 Configurations

// Hardware

GB200 NVL72GB200 Superchip
Configuration36x Grace CPU, 72x B200 GPU1x Grace CPU, 2x B200 GPU
FP4 Tensor Core*1,440 PFLOPS40 PFLOPS
FP8 / FP6 Tensor Core*720 PFLOPS20 PFLOPS
INT8 Tensor Core*720 POPS20 POPS
FP16 / BF16 Tensor Core*360 PFLOPS10 PFLOPS
TF32 Tensor Core*180 PFLOPS5 PFLOPS
FP64 Tensor Core3,240 TFLOPS90 TFLOPS
GPU MemoryUp to 13.5 TB HBM3e, 576 TBpsUp to 384 GB HBM3e, 16 TBps
NVLink Bandwidth130 TBps3.6 TBps
CPU Cores2,952 Arm Neoverse V2 Cores72 Arm Neoverse V2 Cores
CPU MemoryUp to 17 TB LPDDR5X, Up to 18.4 TBpsUp to 480 GB LPDDR5X, Up to 18.4 TB/s
Product infoDatasheet

// Enterprise scale solutions

Supercomputing for any business with ease

Designed for ultra-high-throughput environments, the NVIDIA GB200 NVL72 supports next-generation networking capabilities with bandwidth reaching up to 800 Gb/s.

To maximize AI throughput and eliminate performance bottlenecks, it integrates seamlessly with the latest NVIDIA Quantum-X800 InfiniBand and Spectrum™-X800 Ethernet fabrics.

Furthermore, embedded NVIDIA BlueField-3 DPUs power hyper-scalable AI environments by enabling elastic GPU computing, zero-trust security frameworks, composable storage architectures, and optimized cloud networking.

NVIDIA GB200 NVL72

Lorem ipsum dolor sit amet, consectetur adipiscing elit. Ut elit tellus, luctus nec ullamcorper mattis, pulvinar dapibus leo.

Resources

Continue Exploring

 

Technical walkthrough on how to get started with NVIDIA Quantum-X800.

NVIDIA Silicon Photonics Resources

NVIDIA DGX H200

MAMA AI mDataChat

AI server 8 GPU

AI server 4 GPU

Related Products

Your email address will not be published. Required fields are marked *

Ready For AI Journey?