Transcription of Abstract - arXiv
{{id}} {{{paragraph}}}
Horovod: fast and easy distributed deep learning inTensorFlowAlexander SergeevUber Technologies, Del BalsoUber Technologies, modern deep learning models requires large amounts of computation,often provided by GPUs. Scaling computation from one GPU to many can enablemuch faster training and research progress but entails two complications. First,the training library must support inter-GPU communication. Depending on theparticular methods employed, this communication may entail anywhere fromnegligible to significant overhead. Second, the user must modify his or her trainingcode to take advantage of inter-GPU communication. Depending on the traininglibrary s API, the modification required may be either significant or methods for enabling multi-GPU training under the TensorFlow libraryentail non-negligible communication overhead and require users to heavily mod-ify their model-building code, leading many researchers to avoid the wholemess and stick with slower single-GPU training.
Horovod: fast and easy distributed deep learning in TensorFlow Alexander Sergeev Uber Technologies, Inc. asergeev@uber.com Mike Del Balso Uber Technologies, Inc.
Domain:
Source:
Link to this page:
Please notify us if you found a problem with this document:
{{id}} {{{paragraph}}}
Available in: CUSTOM BUILT | HYBRID LITE |, Available in: CUSTOM BUILT | HYBRID LITE | Standard, CustOM, Build Your Own DI Box, Custom Built, BEST Built-In Ventilation Selection Guide, BEST® Built-In Ventilation Selection Guide, Redi-Built Custom Homes, Built, TANKS FOR SOFTAIL-STYLE & RIGID FRAMES, Open-Web Trusses, Fat Man Fabrication