Dynamic Memory Management for GPU-Based Training of Deep Neural Networks

Shriram S.B,Purushottam Kulkarni,Anshuj Garg

doi:10.1109/ipdps.2019.00030

Abstract

Deep learning has been widely adopted for different applications of artificial intelligence - speech recognition, natural language processing, computer vision etc. The growing size of Deep Neural Networks (DNNs) has compelled the researchers to design memory efficient and performance optimal algorithms. Apart from algorithmic improvements, specialized hardware like Graphics Processing Units (GPUs) are being widely employed to accelerate the training and inference phases of deep networks. However, the limited GPU memory capacity limits the upper bound on the size of networks that can be offloaded to and trained using GPUs. vDNN addresses the GPU memory bottleneck issue and provides a solution which enables training of deep networks that are larger than GPU memory. In our work, we characterize and identify multiple bottlenecks with vDNN like delayed computation start, high pinned memory requirements and GPU memory fragmentation. We present vDNN++ which extends vDNN and resolves the identified issues. Our results show that the performance of vDNN++ is comparable or better (up to 60% relative improvement) than vDNN. We propose different heuristics and order for memory allocation, and empirically evaluate the extent of memory fragmentation with them. We are also able to reduce the pinned memory requirement by up to 60%.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Dynamic Memory Management for GPU-Based Training of Deep Neural Networks

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Acceleration of Large Deep Learning Training with Hybrid GPU Memory Management of Swapping and Re-computing
Haruki Imai ... Tung D Le
-
Haruki Imai, et. al.Haruki Imai ... Tung D Le
10 Dec 2020
10 Dec 2020

MoDNN: Memory Optimal Deep Neural Network Training on Graphics Processing Units
Xiaoming Chen ... Danny Ziyi Chen
IEEE Transactions on Parallel and Distributed Systems | VOL. 30
Xiaoming Chen, et. al.Xiaoming Chen ... Danny Ziyi Chen
01 Mar 2019
IEEE Transactions on Parallel and Distributed Systems | VOL. 30

Ballooning Graphics Memory Space in Full GPU Virtualization Environments
Younghun Park ... Minwoo Gu
Scientific Programming | VOL. 2019
Younghun Park, et. al.Younghun Park ... Minwoo Gu
23 Apr 2019
Scientific Programming | VOL. 2019

MoDNN: Memory optimal DNN training on GPUs
Xiaoming Chen ... Xiaobo Sharon Hu
-
Xiaoming Chen, et. al.Xiaoming Chen ... Xiaobo Sharon Hu
01 Mar 2018
01 Mar 2018

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Dynamic Memory Management for GPU-Based Training of Deep Neural Networks

Abstract

Talk to us

Similar Papers