H100 vs GB200 NVL72 Training Benchmarks – Power, TCO, and Reliability Analysis, Software Improvement Over Time – SemiAnalysis
Frontier model training has pushed GPUs and AI systems to their absolute limits, making cost, efficiency, power, performance per TCO, and reliability central to the discussion on effective training. The Hopper vs Blackwell comparisons are not as simple as Nvidia would have you believe. In this report, we will start by present the results of benchmark runs across over 2,000 H100 GPUs, analyzing data on model flops utilization (MFU), total cost of ownership (TCO) and cost per training 1M tokens. We will also discuss energy use, examining the energy in utility Joules consumed for each token trained and compare it to the average US household annual energy usage, reframing power efficiency in societal context. We will also show the results of this analysis when scaling the GPU cluster from 128 H100s to 2048 H100s and across different versions of Nvidia software. Later in this report, we will also analyze GB200 NVL72 benchmark results across Llama4 400B MoE and DeepSeek 670B MoE and compare
H100 vs GB200 NVL72 Training Benchmarks – Power, TCO, and Reliability Analysis, Software Improvement Over Time – SemiAnalysis Skip to content August 20, 2025 H100 vs GB200 NVL72 Training Benchmarks – Power, TCO, and Reliability Analysis, Software Improvement Over Time Joules per Token, TCO Per Million Tokens, MFU, Tokens Per US Annual Household Energy Usage, DeepSeek 670B, GB200 Unreliability, Backplane Downtime 20 minutes No comments on H100 vs GB200 NVL72 Training Benchmarks – Power, TCO, and Reliability Analysis, Software Improvement Over Time By Dylan Patel and Dani
Explore this link on the map →saved by
related reading
- CVPR2023_eff_tutorial_molchanov.pdfnvlabs.github.io
- GPU Performance Background User's Guide - NVIDIA Docsdocs.nvidia.com
- How To Scale Your Modeljax-ml.github.io
- The Best GPUs for Deep Learning in 2023 — An In-depth Analysistimdettmers.com
- How to Think About GPUs | How To Scale Your Modeljax-ml.github.io
- $2 H100s: How the GPU Bubble Burst - by Eugene Cheahlatent.space
- The Short Case for Nvidia Stock | YouTube Transcript Optimizeryoutubetranscriptoptimizer.com
- Unlocking the full power of NVIDIA H100 GPUs for ML inference with TensorRTbaseten.co
- Nvidia H100 GPUs: Supply and Demand · GPU Utils ⚡️gpus.llm-utils.org
- The Best GPUs for Deep Learning in 2023 — An In-depth Analysistimdettmers.com
- The Scaling Hypothesis · Gwern.netgwern.net
- Reiner Pope – The math behind how LLMs are trained and serveddwarkesh.com