AI Rack-scale Rack-scale systems
Optimized for large-scale AI deployments, AI factories, and converged HPC/AI workloads.
Operationalizing AI at extreme scale
HPE announces NVIDIA Vera Rubin NVL72 by HPE, optimized for AI frontier models over 1 trillion parameters, backed by HPE Services and liquid-cooling expertise.
Unlocking the future of converged HPC/AI workloads
HPE announces NVIDIA GB200 NVL4 by HPE, delivering exceptional performance and density, with integrated HPE management and services.
Fast time to market
Deploy rapidly anywhere in the world with proven ecosystem of HPE Services.
Efficiency
Increase energy efficiency, simplify management and operations with purpose-built, integrated solutions.
Scalability
Expand resources as data and model sizes grow, without disruptive infrastructure changes.
Our customers
Purpose-built portfolio of rack-scale systems for large AI environments
Integrated hardware, software and services to accelerate time to value and simplify deployment.
Our partners and clients
Related products
News and resources
HPE AI Rack-Scale Systems FAQs
What is the portfolio of HPE AI rack-scale systems?
HPE offers a portfolio of rack-scale systems designed for large-scale AI deployments, AI factories, AI data centers, and converged HPC and AI workloads. The portfolio includes NVIDIA GB200 NVL72 by HPE, NVIDIA GB300 NVL72 by HPE, NVIDIA Vera Rubin NVL72 by HPE, NVIDIA GB200 NVL4 by HPE, and AMD Helios AI Rack by HPE. Each system integrates accelerated compute, networking, software, cooling, and HPE Services to simplify the deployment and operation of demanding AI environments.
What customer needs does the HPE AI rack-scale portfolio address?
The portfolio helps organizations deploy and scale AI infrastructure for large model training, fine-tuning, real-time inference, and converged HPC and AI workloads. HPE combines factory-integrated systems with direct liquid cooling, global deployment expertise, performance optimization, and lifecycle support. Customers can accelerate time to value, improve energy efficiency, expand capacity as data and models grow, and operate complex, high-density AI environments more predictably.
When should customers consider NVIDIA GB200 NVL72 by HPE?
NVIDIA GB200 NVL72 by HPE is designed for training, fine-tuning, and inferencing very large AI models, including models with more than one trillion parameters. Its 72 NVIDIA Blackwell GPUs and 36 NVIDIA Grace CPUs are connected through NVIDIA NVLink to provide a single, high-bandwidth 72-GPU domain. The system is well suited for AI service providers and ambitious model builders that need extreme performance and HPE expertise for deployment and ongoing operations.
How is NVIDIA GB300 NVL72 by HPE positioned?
NVIDIA GB300 NVL72 by HPE advances the NVIDIA NVL72 platform with 72 NVIDIA Blackwell Ultra GPUs, 36 NVIDIA Grace CPUs, high-speed NVIDIA NVLink connectivity, and increased GPU memory. It is optimized for massive real-time AI training and inference for models exceeding one trillion parameters. Customers gain a fully integrated, liquid-cooled rack-scale solution supported by HPE deployment, performance engineering, and global services expertise.
What is NVIDIA Vera Rubin NVL72 by HPE designed to deliver?
NVIDIA Vera Rubin NVL72 by HPE is the flagship rack-scale system for frontier AI models, advanced reasoning, and AI agents. It combines 72 NVIDIA Rubin GPUs, 36 NVIDIA Vera CPUs, next-generation NVIDIA NVLink, networking, software, and direct liquid cooling in an integrated design. It targets neoclouds, service providers, leading research organizations, and enterprises that need extreme performance for large AI training and high-volume, low-latency inference at gigascale.
When is NVIDIA GB200 NVL4 by HPE the right choice?
NVIDIA GB200 NVL4 by HPE is optimized for converged HPC and AI workloads that require both GPU acceleration and double-precision performance. Its dense, modular design supports scientific simulation and AI model training, with up to 136 NVIDIA Blackwell GPUs per rack. It is well suited to research labs, universities, and sovereign organizations seeking an Arm-based environment, direct liquid cooling, and HPE Performance Cluster Manager for streamlined cluster deployment and lifecycle management.
What differentiates AMD Helios AI Rack by HPE in the HPE AI rack-scale systems portfolio?
AMD Helios AI Rack by HPE is a high-density system for trillion-parameter model training and high-volume inference, built in an open-standards architecture. It integrates 72 AMD Instinct™ MI455X GPUs with AMD EPYC™ "Venice" CPUs, AMD Pensando™ networking and AMD ROCm™ software in a unified design optimized for power, cooling, and serviceability, while supporting open, interoperable rack-scale fabrics such as UALink™ over Ethernet (UALoE). HPE adds a purpose-built HPE Juniper Networking scale-up Ethernet switch, scale-out switches and extensive liquid-cooling and deployment expertise, giving cloud service providers and neoclouds a flexible alternative designed to reduce proprietary lock-in.
What differentiates HPE from other AI rack-scale system providers?
HPE differentiation extends beyond supplying infrastructure. HPE brings decades of experience designing, integrating, deploying, and supporting the world's largest supercomputing and direct-liquid-cooled environments. Customers can work with one partner for data center design, system integration, installation, validation, performance tuning, resident expertise, global support, and financing. This end-to-end approach helps organizations reduce facility and execution risk while keeping complex AI environments operating predictably throughout their lifecycle.
Which customers are using HPE AI rack-scale systems?
KDDI in Japan is working with HPE to launch AI data center operations using an HPE-built NVIDIA GB200 NVL72 rack-scale system, supporting enterprises and start-ups developing AI applications and training large language models. HLRS selected HPE to build the HammerHAI supercomputer. This system is part of the EU's AI Factory initiatives and is based on NVIDIA GB200 NVL4 by HPE.