paper-with-me

홈 › Papers

Compressed Learning of Deep Neural Networks for OpenCL-Capable Embedded Systems

2019-05-20 · Sangkyun Lee, Jeonghyun Lee

Deep neural networks (DNNs) have been quite successful in solving many complex learning problems. However, DNNs tend to have a large number of learning parameters, leading to a large memory and computation requirement. In this paper, we propose a model compression framework for efficient training and inference of deep neural networks on embedded systems. Our framework provides data structures and kernels for OpenCL-based parallel forward and backward computation in a compressed form. In particular, our method learns sparse representations of parameters using $\ell_1$-based sparse coding while training, storing them in compressed sparse matrices. Unlike the previous works, our method does not require a pre-trained model as an input and therefore can be more versatile for different application environments. Even though the use of $\ell_1$-based sparse coding for model compression is not new, we show that it can be far more effective than previously reported when we use proximal point algorithms and the technique of debiasing. Our experiments show that our method can produce minimal learning models suitable for small embedded devices.

📄 PDF Abstract BibTeX arXiv:1905.07931

Code (1)

sanglee/caffe-mc-opencl 공식 구현

Tasks

Model Compression

Similar Papers 제목 키워드 기반

OpenClinicalAI: enabling AI to diagnose diseases in real-world clinical settings

2021-09-09 · Yunyou Huang, Nana Wang, Suqin Tang, Li Ma 외

This paper quantitatively reveals the state-of-the-art and state-of-the-practice AI systems only achieve acceptable performance on the stringent conditions that all categories of subjects are known, which we call closed …

CLBlast: A Tuned OpenCL BLAS Library

2017-05-12 · Cedric Nugteren

This work introduces CLBlast, an open-source BLAS library providing optimized OpenCL routines to accelerate dense linear algebra for a wide variety of devices. It is targeted at machine learning and HPC applications and …

Toward a Modular Architecture for Embedded AI Agent Systems at the Edge

2026-06-01 · Marcus Rüb, Michael Gerhards arxiv

The rise of Large Language Models (LLMs) has enabled agentic AI capable of complex reasoning and tool use; however, deploying such autonomy in pervasive computing environments remains challenging due to the strict memory…

Security of OpenClaw Agents: Fundamentals, Attacks, and Countermeasures

2026-05-25 · Yuntao Wang, Jianle Ba, Han Liu, Yanghe Pan 외 arxiv

The rapid evolution of large language model (LLM)-driven autonomous agents has given rise to OpenClaw, a new class of open-source agent frameworks that operate as continuously running, skill-augmented systems with persis…

From Chatbot to Digital Colleague: The Paradigm Shift Toward Persistent Autonomous AI

2026-06-12 · Yongheng Zhang, Ziang Liu, Jiaxuan Zhu, Shuai Wang 외 arxiv

Large Language Models (LLMs) are undergoing a fundamental transformation from conversational generators into integrated AI systems capable of reasoning, action, memory, and self-improvement. We conceptualize this transit…

Reinforcement Learning