arXiv · 2603.27831
Quantifying and Attributing Power Flexibility from GPU-Heavy Data Centers
Abstract
The growth of GPU-heavy data centers has increased electricity demand and challenged grid stability. This paper investigates how an energy-aware job scheduling algorithm provides flexibility in GPU-heavy data centers. We develop a rolling-horizon optimization framework considering IT power and cooling dynamics with limited future job information. Compared with the first-in first-out baseline, we show that energy-aware scheduling brings latent power flexibility during peak-price periods. This flexibility is created through both thermal and computational mechanisms: cooling shifting can reliably reduce demand for short periods at relatively low incentive (\$30/MWh), and movement of backfilled jobs can often reduce demand at similar prices (\$30-300/MWh). Further reduction is possible through reordering or delaying jobs, but due to lost profits these actions come at higher prices (starting at \$600/MWh, more significantly above \$3000/MWh). Flexibility is achievable without knowing arriving jobs, but much greater flexibility can be achieved with perfect foresight of the future queue.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Yiru Ji, Constance Crozier, Matthew Liska. 2026-03-29. Quantifying and Attributing Power Flexibility from GPU-Heavy Data Centers. https://arxiv.org/abs/2603.27831
Cite the original work for its findings. Save a collection to share your selection of sources.