paper-with-me

홈 › Papers

Caffe con Troll: Shallow Ideas to Speed Up Deep Learning

2015-04-16 · Stefan Hadjis, Firas Abuzaid, Ce Zhang, Christopher Ré

We present Caffe con Troll (CcT), a fully compatible end-to-end version of the popular framework Caffe with rebuilt internals. We built CcT to examine the performance characteristics of training and deploying general-purpose convolutional neural networks across different hardware architectures. We find that, by employing standard batching optimizations for CPU training, we achieve a 4.5x throughput improvement over Caffe on popular networks like CaffeNet. Moreover, with these improvements, the end-to-end training time for CNNs is directly proportional to the FLOPS delivered by the CPU, which enables us to efficiently train hybrid CPU-GPU systems for CNNs.

📄 PDF Abstract BibTeX arXiv:1504.04343

Code (1)

HazyResearch/CaffeConTroll 공식 구현

Tasks

CPUDeep LearningGPU

Similar Papers 제목 키워드 기반

HG-Caffe: Mobile and Embedded Neural Network GPU (OpenCL) Inference Engine with FP16 Supporting

2019-01-03 · Zhuoran Ji

Breakthroughs in the fields of deep learning and mobile system-on-chips are radically changing the way we use our smartphones. However, deep neural networks inference is still a challenging task for edge AI devices due t…

GPU

Using GPI-2 for Distributed Memory Paralleliziation of the Caffe Toolbox to Speed up Deep Neural Network Training

2017-05-31 · Martin Kuehn, Janis Keuper, Franz-Josef Pfreundt

Deep Neural Network (DNN) are currently of great inter- est in research and application. The training of these net- works is a compute intensive and time consuming task. To reduce training times to a bearable amount at r…

Blockingvalid

MyCaffe: A Complete C# Re-Write of Caffe with Reinforcement Learning

2018-10-04 · David W. Brown

Over the past few years Caffe, from Berkeley AI Research, has gained a strong following in the deep learning community with over 15K forks on the github.com/BLVC/Caffe site. With its well organized, very modular C++ desi…

Deep Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Tuning of Mixture-of-Experts Mixed-Precision Neural Networks

2022-09-29 · Fabian Tschopp

Deep learning has become a useful data analysis method, however mainstream adaption in distributed computer software and embedded devices has been low so far. Often, adding deep learning inference in mainstream applicati…

image-classificationImage ClassificationMixture-of-Experts

PyText: A Seamless Path from NLP research to production

2018-12-12 · Ahmed Aly, Kushal Lakhotia, Shicong Zhao, Mrinal Mohit 외

We introduce PyText - a deep learning based NLP modeling framework built on PyTorch. PyText addresses the often-conflicting requirements of enabling rapid experimentation and of serving models at scale. It achieves this …