Liquid-Cooled AI SuperCluster

Liquid-Cooled AI SuperCluster

With 256 NVIDIA HGX™ B200 GPUs, 32 4U Liquid-cooled Systems

Liquid-cooled AI SuperCluster
DON'T OVERTHINK — SEE PRODUCTS GET PRICING

Unprecedented Density and Efficiency

These new SuperCluster offerings powered by the NVIDIA Blackwell Platform are available in 42U, 48U, or 52U configurations. The upgraded cold plates and 250kW coolant distribution unit (CDU) more than double the cooling capacity of the previous generation. The new vertical coolant distribution manifold (CDM) means that horizontal manifolds no longer occupy valuable rack space. NVIDIA Quantum™ InfiniBand or Spectrum™-4 400GbE networking in a centralized rack enables a non-blocking, 256-GPU scalable unit in five racks, or an extended 768-GPU scalable unit in nine racks.

Software & Services

Software: Supermicro’s SuperCloud Composer software provides management tools for monitoring and optimizing air or liquid-cooled infrastructure, delivering a complete solution from proof of concept to full-scale deployment. Manage all data center racks, including compute, storage, and networking in one unified dashboard.

SuperCluster natively supports NVIDIA AI Enterprise software to accelerate time to online for production AI. NVIDIA NIM microservices allow organizations to easily access and deploy the latest AI models and AI agents, fully optimized for the new NVIDIA Blackwell Platforms.

Services: Supermicro’s on-site rack deployment helps enterprises build a data center from the ground up, including planning, design, power-up, validation, testing, installation, and configuration of racks, servers, switches, and other networking equipment to meet the organization’s specific needs.

Supermicro NVIDIA HGX B200 8-GPU Systems, Liquid-Cooled

Supermicro NVIDIA HGX Systems power the world’s largest liquid-cooled AI data centers. The new 4U NVIDIA HGX B200 8-GPU system features new cold plates and tubing design that further enhances efficiency and serviceability over its predecessor. The system features 8 NVIDIA Blackwell GPUs, each with 180 GB HBM3e memory. The GPUs are interconnected at 1.8 TB/s through the latest NVIDIA NVLink, with 1.4 TB of GPU memory capacity per system.

The SuperCluster creates a massive pool of GPU resources, acting as one AI supercomputer, featuring 1:1 networking to GPU with 8× 400 Gb/s NVIDIA ConnectX-7 adapters or BlueField-3 SuperNICs, as well as 2 NVIDIA BlueField-3 DPUs per system.

Node Configuration
Overview4U Liquid-cooled system with NVIDIA HGX B200 8-GPU
CPUDual Intel® Xeon® 6900 series (SYS-422GA-NBRT-LCC)
Dual AMD EPYC™ 9005/9004 (AS-4126GS-NBRT-LCC)
Dual 5th/4th Gen Intel® Xeon® Scalable (SYS-421GE-NBRT-LCC)
Memory24 DIMMs up to DDR5-6400
24 DIMMs up to DDR5-6000
32 DIMMs up to DDR5-5600
GPU8× NVIDIA HGX B200 (180 GB HBM3e each)
1.8 TB/s NVIDIA NVLink with NVSwitch
Networking8× single-port NVIDIA ConnectX-7 or BlueField-3 SuperNICs
2× dual-port BlueField-3 DPUs
Storage8× hot-swap 2.5” NVMe bays
2× M.2 NVMe slots
Power Supply4× 6600 W redundant Titanium PSUs

*Recommended configuration; other memory, networking, storage options available.

Node Configuration

Scalable Units

32-Node Scalable Unit

32-Node Scalable Unit
OverviewFully integrated liquid-cooled 32-node cluster with 256 NVIDIA B200 GPUs
Compute Fabric8× NVIDIA Quantum-2 400G InfiniBand or Spectrum-4 400GbE switches
In-band Mgmt3× Spectrum SN4600 100GbE switches
Out-of-band Mgmt2× SSE-G3748R-SMIS 48-port 1GbE ToR
1× SSE-F3548SR 48-port 10GbE ToR
Liquid Cooling4× Supermicro 250 kW CDUs with redundant PSU & pumps

96-Node Scalable Unit

96-Node Scalable Unit
OverviewFully integrated liquid-cooled 96-node cluster with 768 NVIDIA B200 GPUs
Compute Fabric24× NVIDIA Quantum-2 400G InfiniBand or Spectrum-4 400GbE switches
In-band Mgmt9× Spectrum SN4600 100GbE switches
Out-of-band Mgmt6× SSE-G3748R-SMIS 48-port 1GbE ToR
3× SSE-F3548SR 48-port 10GbE ToR
Liquid Cooling8× Supermicro 250 kW CDUs with redundant PSU & pumps

Rack Scale Design Close-up

Rack Scale Design Close-up

Networking

  • NVIDIA Quantum-2 400G InfiniBand or Spectrum-4 400GbE switches for compute & storage
  • Ethernet leaf switches for in-band management
  • Out-of-band 1G/10G IPMI switch
  • Non-blocking network topology

Compute

  • 8× SYS-422GA-NBRT-LCC / AS-4126GS-NBRT-LCC / SYS-421GE-NBRT-LCC per rack
  • 64× NVIDIA HGX B200 GPUs per rack
  • 11.5 TB HBM3e per rack
  • Flexible storage with NVIDIA GPUDirect RDMA & RoCE support

Liquid-Cooling

  • Supermicro 250 kW Cooling Distribution Unit (CDU) with redundant PSU & hot-swap pumps
  • Vertical Cooling Distribution Manifold (CDM)
View as Grid List

6 items available

Set Descending Direction
Products
  1. GPU SuperServer SYS-422GA-NBRT-LCC SYS-422GA-NBRT-LCC Supermicro GPU SuperServer

    Scientific Research
    Conversational AI
    Business Intelligence & Analytics
    Drug Discovery
    Finance Services and Fraud Detection
    AI/Deep Learning Training and Inference
    Large Language Model (LLM) and Generative AI
    High Performance Computing (HPC)
    Autonomous Vehicle Technologies

    413 398.69 €
per page
Supporting Products
Contact us to learn more about our solutions
Contact now

ServerSimply Liquid Cooled AI SuperCluster Solutions

Step into the next generation of AI compute performance with ServerSimply Liquid Cooled AI SuperCluster Solutions, featuring advanced liquid-cooling systems that ensure optimal thermal efficiency for high-density NVIDIA® HGX GPU clusters. Our solution supports both Ethernet and InfiniBand fabrics, delivering ultra-low latency and maximum throughput for demanding AI and HPC workloads. Designed for scalability and reliability, ServerSimply’s Liquid Cooled AI SuperCluster enables seamless expansion from single-rack deployments to exascale configurations, providing businesses with a future-proof platform that maximizes performance, reduces operational costs, and maintains superior uptime.

Loading...