paper-with-me

Papers

ChainerMN: Scalable Distributed Deep Learning Framework

2017-10-31 · Takuya Akiba, Keisuke Fukuda, Shuji Suzuki

One of the keys for deep learning to have made a breakthrough in various fields was to utilize high computing powers centering around GPUs. Enabling the use of further computing abilities by distributed processing is essential not only to make the deep learning bigger and faster but also to tackle unsolved challenges. We present the design, implementation, and evaluation of ChainerMN, the distributed deep learning framework we have developed. We demonstrate that ChainerMN can scale the learning process of the ResNet-50 model to the ImageNet dataset up to 128 GPUs with the parallel efficiency of 90%.

📄 PDF Abstract BibTeX arXiv:1710.11351

Code (1)

chainer/chainermn 공식 구현

Tasks

Deep Learning

Similar Papers 제목 키워드 기반

Evolving Large-Scale Data Stream Analytics based on Scalable PANFIS

2018-07-18 · Mahardhika Pratama, Choiru Za'in, Eric Pardede

Many distributed machine learning frameworks have recently been built to speed up the large-scale data learning process. However, most distributed machine learning used in these frameworks still uses an offline algorithm…

Active Learning

DRASIC: Distributed Recurrent Autoencoder for Scalable Image Compression

2019-03-23 · Enmao Diao, Jie Ding, Vahid Tarokh

We propose a new architecture for distributed image compression from a group of distributed data sources. The work is motivated by practical needs of data-driven codec design, low power consumption, robustness, and data …

DecoderImage Compression

SHADHO: Massively Scalable Hardware-Aware Distributed Hyperparameter Optimization

2017-07-05 · Jeff Kinnison, Nathaniel Kremer-Herman, Douglas Thain, Walter Scheirer

Computer vision is experiencing an AI renaissance, in which machine learning models are expediting important breakthroughs in academic research and commercial applications. Effectively training these models, however, is …

Cell SegmentationHyperparameter Optimization

Scalable AI-assisted Workflow Management for Detector Design Optimization Using Distributed Computing

2026-03-31 · Derek Anderson, Amit Bashyal, Markus Diefenthaler, Cristiano Fanelli 외 arxiv

The Production and Distributed Analysis (PanDA) system, originally developed for the ATLAS experiment at the CERN Large Hadron Collider (LHC), has evolved into a robust platform for orchestrating large-scale workflows ac…

Energy-Harvesting Distributed Machine Learning

2021-02-10 · Basak Guler, Aylin Yener

This paper provides a first study of utilizing energy harvesting for sustainable machine learning in distributed networks. We consider a distributed learning setup in which a machine learning model is trained over a larg…

BIG-bench Machine LearningEdge-computing