paper-with-me

Papers

Scheduling Real-time Deep Learning Services as Imprecise Computations

2020-11-02 · Shuochao Yao, Yifan Hao, Yiran Zhao, Huajie Shao, Dongxin Liu, Shengzhong Liu, Tianshi Wang, Jinyang Li, Tarek Abdelzaher

The paper presents an efficient real-time scheduling algorithm for intelligent real-time edge services, defined as those that perform machine intelligence tasks, such as voice recognition, LIDAR processing, or machine vision, on behalf of local embedded devices that are themselves unable to support extensive computations. The work contributes to a recent direction in real-time computing that develops scheduling algorithms for machine intelligence tasks with anytime prediction. We show that deep neural network workflows can be cast as imprecise computations, each with a mandatory part and (several) optional parts whose execution utility depends on input data. The goal of the real-time scheduler is to maximize the average accuracy of deep neural network outputs while meeting task deadlines, thanks to opportunistic shedding of the least necessary optional parts. The work is motivated by the proliferation of increasingly ubiquitous but resource-constrained embedded devices (for applications ranging from autonomous cars to the Internet of Things) and the desire to develop services that endow them with intelligence. Experiments on recent GPU hardware and a state of the art deep neural network for machine vision illustrate that our scheme can increase the overall accuracy by 10%-20% while incurring (nearly) no deadline misses.

📄 PDF Abstract BibTeX arXiv:2011.01112

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningGPUScheduling

Similar Papers 제목 키워드 기반

FSMoE: A Flexible and Scalable Training System for Sparse Mixture-of-Experts Models

2025-01-18 · Xinglin Pan, WenXiang Lin, Lin Zhang, Shaohuai Shi 외

Recent large language models (LLMs) have tended to leverage sparsity to reduce computations, employing the sparsely activated mixture-of-experts (MoE) technique. MoE introduces four modules, including token routing, toke…

GPUMixture-of-ExpertsScheduling

A proof of contribution in blockchain using game theoretical deep learning model

2024-08-25 · Jin Wang

Building elastic and scalable edge resources is an inevitable prerequisite for providing platform-based smart city services. Smart city services are delivered through edge computing to provide low-latency applications. H…

Edge-computingGraph Neural NetworkScheduling

Intelligent Resource Scheduling for Co-located Latency-critical Services: A Multi-Model Collaborative Learning Approach

2019-11-26 · Lei Liu

Latency-critical services have been widely deployed in cloud environments. For cost-efficiency, multiple services are usually co-located on a server. Thus, run-time resource scheduling becomes the pivot for QoS control i…

BIG-bench Machine LearningReinforcement LearningScheduling

Zygarde: Time-Sensitive On-Device Deep Inference and Adaptation on Intermittently-Powered Systems

2019-05-05 · Bashima Islam, Shahriar Nirjon

We propose Zygarde -- which is an energy -- and accuracy-aware soft real-time task scheduling framework for batteryless systems that flexibly execute deep learning tasks1 that are suitable for running on microcontrollers…

Scheduling

Improving non-deterministic uncertainty modelling in Industry 4.0 scheduling

2021-01-08 · Ashwin Misra, Ankit Mittal, Vihaan Misra, Deepanshu Pandey

The latest Industrial revolution has helped industries in achieving very high rates of productivity and efficiency. It has introduced data aggregation and cyber-physical systems to optimize planning and scheduling. Altho…

ArticlesDecision MakingScheduling