Hewlett Packard Enterprise Accelerates AI Training with New Turnkey Solution Powered by NVIDIA
Summary
Using HPE Machine Learning Development Environment on this system, the open source 70 billion-parameter Llama 2 model was fine-tuned in less than 3 minutesi, translating directly to faster time-to-value for customers. • Turnkey simplicity – The solution is complemented by HPE Complete Care Services which provides global specialists for set-up, installation and full lifecycle support to simplify AI adoption. Built on decades of reimagining the future and innovating to advance the way people live and work, HPE delivers unique, open and intelligent technology solutions as a service. For more information, visit: www.hpe.com i Using 32 HPE Cray EX 2500 nodes with 128 NVIDIA H100 GPUs at 97% scaling efficiency, a 70 billion-parameter Llama 2 model was fine-tuned in internal tests on a 10 million token corpus in less than 3 minutes. The independently-run tests showed 2-3X performance improvement as compared to MLPerf 3.0 published results for an A100-based system comprising two AMD EPYC 7763 processors and four NVIDIA A100 GPUs with NVLINK interconnects.