arXiv · 2603.27831
Quantifying and Attributing Power Flexibility from GPU-Heavy Data Centers
Abstract
The growth of GPU-heavy data centers has increased electricity demand and challenged grid stability. This paper investigates how an energy-aware job scheduling algorithm provides flexibility in GPU-heavy data centers. We develop a rolling-horizon optimization framework considering IT power and cooling dynamics with limited future job information. Compared with the first-in first-out baseline, we show that energy-aware scheduling brings latent power flexibility during peak-price periods. This flexibility is created through both thermal and computational mechanisms: cooling shifting can reliably reduce demand for short periods at relatively low incentive (\$30/MWh), and movement of backfilled jobs can often reduce demand at similar prices (\$30-300/MWh). Further reduction is possible through reordering or delaying jobs, but due to lost profits these actions come at higher prices (starting at \$600/MWh, more significantly above \$3000/MWh). Flexibility is achievable without knowing arriving jobs, but much greater flexibility can be achieved with perfect foresight of the future queue.
Explore related subjects
Keep this discovery
Yiru Ji, Constance Crozier, Matthew Liska. 2026-03-29. Quantifying and Attributing Power Flexibility from GPU-Heavy Data Centers. https://arxiv.org/abs/2603.27831
Cite the original work for its findings. Save a collection to share your selection of sources.