paper-with-me

홈 › Papers

Neural Network Inference on Mobile SoCs

2019-08-24 · Siqi Wang, Anuj Pathania, Tulika Mitra

The ever-increasing demand from mobile Machine Learning (ML) applications calls for evermore powerful on-chip computing resources. Mobile devices are empowered with heterogeneous multi-processor Systems-on-Chips (SoCs) to process ML workloads such as Convolutional Neural Network (CNN) inference. Mobile SoCs house several different types of ML capable components on-die, such as CPU, GPU, and accelerators. These different components are capable of independently performing inference but with very different power-performance characteristics. In this article, we provide a quantitative evaluation of the inference capabilities of the different components on mobile SoCs. We also present insights behind their respective power-performance behavior. Finally, we explore the performance limit of the mobile SoCs by synergistically engaging all the components concurrently. We observe that a mobile SoC provides up to 2x improvement with parallel inference when all its components are engaged, as opposed to engaging only one component.

📄 PDF Abstract BibTeX arXiv:1908.11450

Code (0)

등록된 구현이 없습니다.

Tasks

CPUGPU

Similar Papers 제목 키워드 기반

HeteroLLM: Accelerating Large Language Model Inference on Mobile SoCs platform with Heterogeneous AI Accelerators

2025-01-11 · Le Chen, Dahu Feng, Erhu Feng, Rong Zhao 외

With the rapid advancement of artificial intelligence technologies such as ChatGPT, AI agents and video generation,contemporary mobile systems have begun integrating these AI capabilities on local devices to enhance priv…

Language ModelingLanguage ModellingLarge Language ModelVideo Generation

AI Benchmark: All About Deep Learning on Smartphones in 2019

2019-10-15 · Andrey Ignatov, Radu Timofte, Andrei Kulik, Seungsoo Yang 외

The performance of mobile AI accelerators has been evolving rapidly in the past two years, nearly doubling with each new generation of SoCs. The current 4th generation of mobile NPUs is already approaching the results of…

AllDeep Learning

Understanding Large Language Models in Your Pockets: Performance Study on COTS Mobile Devices

2024-10-04 · Jie Xiao, Qianyi Huang, Xu Chen, Chen Tian

As large language models (LLMs) increasingly integrate into every aspect of our work and daily lives, there are growing concerns about user privacy, which push the trend toward local deployment of these models. There are…

BenchmarkingLanguage ModelingLanguage ModellingLarge Language Model

PIM-AI: A Novel Architecture for High-Efficiency LLM Inference

2024-11-26 · Cristobal Ortega, Yann Falevoz, Renaud Ayrignac

Large Language Models (LLMs) have become essential in a variety of applications due to their advanced language understanding and generation capabilities. However, their computational and memory requirements pose signific…

Real Image Denoising with Knowledge Distillation for High-Performance Mobile NPUs

2026-05-05 · Faraz Kayani, Sarmad Kayani, Asad Ahmed, Radu Timofte 외 arxiv

While deep-learning-based image restoration has achieved unprecedented fidelity, deployment on mobile Neural Processing Units (NPUs) remains bottlenecked by operator incompatibility and memory-access overhead. We propose…

Knowledge DistillationImage RestorationImage Denoising