paper-with-me

홈 › Papers

AutoDiCE: Fully Automated Distributed CNN Inference at the Edge

2022-07-20 · Xiaotian Guo, Andy D. Pimentel, Todor Stefanov

Deep Learning approaches based on Convolutional Neural Networks (CNNs) are extensively utilized and very successful in a wide range of application areas, including image classification and speech recognition. For the execution of trained CNNs, i.e. model inference, we nowadays witness a shift from the Cloud to the Edge. Unfortunately, deploying and inferring large, compute and memory intensive CNNs on edge devices is challenging because these devices typically have limited power budgets and compute/memory resources. One approach to address this challenge is to leverage all available resources across multiple edge devices to deploy and execute a large CNN by properly partitioning the CNN and running each CNN partition on a separate edge device. Although such distribution, deployment, and execution of large CNNs on multiple edge devices is a desirable and beneficial approach, there currently does not exist a design and programming framework that takes a trained CNN model, together with a CNN partitioning specification, and fully automates the CNN model splitting and deployment on multiple edge devices to facilitate distributed CNN inference at the Edge. Therefore, in this paper, we propose a novel framework, called AutoDiCE, for automated splitting of a CNN model into a set of sub-models and automated code generation for distributed and collaborative execution of these sub-models on multiple, possibly heterogeneous, edge devices, while supporting the exploitation of parallelism among and within the edge devices. Our experimental results show that AutoDiCE can deliver distributed CNN inference with reduced energy consumption and memory usage per edge device, and improved overall system throughput at the same time.

📄 PDF Abstract BibTeX arXiv:2207.12113

Code (1)

parrotsky/autodice 공식 구현

Tasks

Code Generationimage-classificationImage Classificationspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

HiDP: Hierarchical DNN Partitioning for Distributed Inference on Heterogeneous Edge Platforms

2024-11-25 · Zain Taufique, Aman Vyas, Antonio Miele, Pasi Liljeberg 외

Edge inference techniques partition and distribute Deep Neural Network (DNN) inference tasks among multiple edge nodes for low latency inference, without considering the core-level heterogeneity of edge nodes. Further, d…

Tessel: Boosting Distributed Execution of Large DNN Models via Flexible Schedule Search

2023-11-26 · Zhiqi Lin, Youshan Miao, Guanbin Xu, Cheng Li 외

Increasingly complex and diverse deep neural network (DNN) models necessitate distributing the execution across multiple devices for training and inference tasks, and also require carefully planned schedules for performa…

Priority-Aware Model-Distributed Inference at Edge Networks

2024-12-16 · Teng Li, Hulya Seferoglu

Distributed inference techniques can be broadly classified into data-distributed and model-distributed schemes. In data-distributed inference (DDI), each worker carries the entire Machine Learning (ML) model but processe…

Fluid Dynamic DNNs for Reliable and Adaptive Distributed Inference on Edge Devices

2024-01-17 · Lei Xun, Mingyu Hu, Hengrui Zhao, Amit Kumar Singh 외

Distributed inference is a popular approach for efficient DNN inference at the edge. However, traditional Static and Dynamic DNNs are not distribution-friendly, causing system reliability and adaptability issues. In this…

New Directions in Distributed Deep Learning: Bringing the Network at Forefront of IoT Design

2020-08-25 · Kartikeya Bhardwaj, Wei Chen, Radu Marculescu

In this paper, we first highlight three major challenges to large-scale adoption of deep learning at the edge: (i) Hardware-constrained IoT devices, (ii) Data security and privacy in the IoT era, and (iii) Lack of networ…

Deep LearningFederated Learning