Hardware
Two classes of Blackwell compute, one data platform.
Tightly coupled DGX systems for the largest training jobs, flexible
high-memory GPU nodes for everything else, and VAST all-flash storage feeding both over a
dedicated fabric. Full specifications are in the
hardware guide.
128 × B200
Flagship AI nodes
16 NVIDIA DGX B200 systems for large-scale model training and tightly
coupled multi-GPU science.
- GPUs per node
- 8× B200, NVSwitch
- GPU memory
- ~180 GB HBM each
- System memory
- 2 TB / node
- CPU
- 112 Xeon cores / node
- Interconnect
- 800 Gb/s NDR IB
- Local NVMe
- 34.2 TB / node
336 × RTX PRO 6000
Flexible AI compute
42 Blackwell RTX nodes for fine-tuning, inference, simulation, and
data-intensive workloads.
- GPUs per node
- 8× RTX PRO 6000
- GPU memory
- 96 GB GDDR7 each
- System memory
- 1–2 TB / node
- CPU
- 128 Xeon cores / node
- Interconnect
- 400 Gb/s ConnectX-7
- Local NVMe
- 31.68 TB / node
10.8 PB flash
Data platform
VAST all-flash storage built for high-bandwidth AI workloads and large
shared datasets.
- Appliances
- 8× VAST Ceres V2
- Raw NVMe
- ~10.8 PB shared
- Acceleration
- 12.8 TB SCM each
- Offload
- BlueField-3 DPUs
- Storage fabric
- 200–400 Gb Ethernet
- Node-local NVMe
- 1.9 PB aggregate
Fabric
NDR InfiniBand at up to 800 Gb/s
Non-blocking 8-rail topology across DGX nodes
Dedicated storage fabric between compute and VAST
BlueField-3 DPUs for accelerated networking
System totals also count 12 management and interactive nodes
(1,152 CPU cores, 6 TB memory) that run system services and user access.