arXiv · 2003.08732
Addressing the Memory Bottleneck in AI Model Training
Abstract
Using medical imaging as case-study, we demonstrate how Intel-optimized TensorFlow on an x86-based server equipped with 2nd Generation Intel Xeon Scalable Processors with large system memory allows for the training of memory-intensive AI/deep-learning models in a scale-up server configuration. We believe our work represents the first training of a deep neural network having large memory footprint (~ 1 TB) on a single-node server. We recommend this configuration to scientists and researchers who wish to develop large, state-of-the-art AI models but are currently limited by memory.
Explore related subjects
Keep this discovery
David Ojika, Bhavesh Patel, G. Anthony Reina, Trent Boyer, Chad Martin, Prashant Shah. 2020-03-11. Addressing the Memory Bottleneck in AI Model Training. https://arxiv.org/abs/2003.08732
Cite the original work for its findings. Save a collection to share your selection of sources.