paper-with-me

Papers

Optimizing edge AI models on HPC systems with the edge in the loop

2025-05-26 · Marcel Aach, Cyril Blanc, Andreas Lintermann, Kurt De Grave

Artificial intelligence and machine learning models deployed on edge devices, e.g., for quality control in Additive Manufacturing (AM), are frequently small in size. Such models usually have to deliver highly accurate results within a short time frame. Methods that are commonly employed in literature start out with larger trained models and try to reduce their memory and latency footprint by structural pruning, knowledge distillation, or quantization. It is, however, also possible to leverage hardware-aware Neural Architecture Search (NAS), an approach that seeks to systematically explore the architecture space to find optimized configurations. In this study, a hardware-aware NAS workflow is introduced that couples an edge device located in Belgium with a powerful High-Performance Computing system in Germany, to train possible architecture candidates as fast as possible while performing real-time latency measurements on the target hardware. The approach is verified on a use case in the AM domain, based on the open RAISE-LPBF dataset, achieving ~8.8 times faster inference speed while simultaneously enhancing model quality by a factor of ~1.35, compared to a human-designed baseline.

📄 PDF Abstract BibTeX arXiv:2505.19995

Code (1)

flanders-make-vzw/hpc2edge 공식 구현 pytorch

Tasks

Hardware Aware Neural Architecture SearchKnowledge DistillationNeural Architecture SearchQuantization

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…
AM 설명 없음

Similar Papers 제목 키워드 기반

Neuromorphic Energy-Aware Learning for Adaptive Deep Brain Stimulation

2026-06-26 · Binh Nguyen, Colleen Josephson, Mircea Teodorescu, Gert Cauwenberghs 외 arxiv

Neuromorphic and edge computing research has focused on reducing the inference cost of neural network controllers, yet in physical closed-loop systems the actuator can rival or exceed an efficient controller in energy. A…

Reinforcement LearningKnowledge Distillation

Data-Driven Optimized Tracking Control Heuristic for MIMO Structures: A Balance System Case Study

2021-04-01 · Ning Wang, Mohammed Abouheaf, Wail Gueaieb

A data-driven computational heuristic is proposed to control MIMO systems without prior knowledge of their dynamics. The heuristic is illustrated on a two-input two-output balance system. It integrates a self-adjusting n…

Integration of Prior Knowledge into Direct Learning for Safe Control of Linear Systems

2025-02-06 · Amir Modares, Bahare Kiumarsi, Hamidreza Modares

This paper integrates prior knowledge into direct learning of safe controllers for linear uncertain systems under disturbances. To this end, we characterize the set of all closed-loop systems that can be explained by ava…

Intelligent Sensing-to-Action for Robust Autonomy at the Edge: Opportunities and Challenges

2025-02-04 · Amit Ranjan Trivedi, Sina Tayebati, Hemant Kumawat, Nastaran Darabi 외

Autonomous edge computing in robotics, smart cities, and autonomous vehicles relies on the seamless integration of sensing, processing, and actuation for real-time decision-making in dynamic environments. At its core is …

Autonomous VehiclesEdge-computing

Edge Graph Intelligence: Reciprocally Empowering Edge Networks with Graph Intelligence

2024-07-07 · Liekang Zeng, Shengyuan Ye, Xu Chen, Xiaoxi Zhang 외

Recent years have witnessed a thriving growth of computing facilities connected at the network edge, cultivating edge networks as a fundamental infrastructure for supporting miscellaneous intelligent services.Meanwhile, …

Edge-computingGraph LearningGraph Representation LearningMiscellaneous+1