paper-with-me

홈 › Papers

Ariel-ML: Computing Parallelization with Embedded Rust for Neural Networks on Heterogeneous Multi-core Microcontrollers

2025-12-10 · Zhaolan Huang, Kaspar Schleiser, Gyungmin Myung, Emmanuel Baccelli arxiv

Low-power microcontroller (MCU) hardware is currently evolving from single-core architectures to predominantly multi-core architectures. In parallel, new embedded software building blocks are more and more written in Rust, while C/C++ dominance fades in this domain. On the other hand, small artificial neural networks (ANN) of various kinds are increasingly deployed in edge AI use cases, thus deployed and executed directly on low-power MCUs. In this context, both incremental improvements and novel innovative services will have to be continuously retrofitted using ANNs execution in software embedded on sensing/actuating systems already deployed in the field. However, there was so far no Rust embedded software platform automating parallelization for inference computation on multi-core MCUs executing arbitrary TinyML models. This paper thus fills this gap by introducing Ariel-ML, a novel toolkit we designed combining a generic TinyML pipeline and an embedded Rust software platform which can take full advantage of multi-core capabilities of various 32bit microcontroller families (Arm Cortex-M, RISC-V, ESP-32). We published the full open source code of its implementation, which we used to benchmark its capabilities using a zoo of various TinyML models. We show that Ariel-ML outperforms prior art in terms of inference latency as expected, and we show that, compared to pre-existing toolkits using embedded C/C++, Ariel-ML achieves comparable memory footprints. Ariel-ML thus provides a useful basis for TinyML practitioners and resource-constrained embedded Rust developers.

📄 PDF Abstract BibTeX arXiv:2512.09800

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Accelerating AI and Computer Vision for Satellite Pose Estimation on the Intel Myriad X Embedded SoC

2024-09-19 · Vasileios Leon, Panagiotis Minaidis, George Lentaris, Dimitrios Soudris

The challenging deployment of Artificial Intelligence (AI) and Computer Vision (CV) algorithms at the edge pushes the community of embedded computing to examine heterogeneous System-on-Chips (SoCs). Such novel computing …

DiversityPose Estimation

Gegelati: Lightweight Artificial Intelligence through Generic and Evolvable Tangled Program Graphs

2020-12-15 · Karol Desnos, Nicolas Sourbier, Pierre-Yves Raumer, Olivier Gesny 외

Tangled Program Graph (TPG) is a reinforcement learning technique based on genetic programming concepts. On state-of-the-art learning environments, TPGs have been shown to offer comparable competence with Deep Neural Net…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Small Language Models as Compiler Experts: Auto-Parallelization for Heterogeneous Systems

2025-12-22 · Prathamesh Devadiga arxiv

Traditional auto-parallelizing compilers, reliant on rigid heuristics, struggle with the complexity of modern heterogeneous systems. This paper presents a comprehensive evaluation of small (approximately 1B parameter) la…

Rethinking Dynamic Networks and Heterogeneous Computing with Automatic Parallelization

2025-06-03 · Ruilong Wu, Xinjiao Li, Yisu Wang, Xinyu Chen 외

Hybrid parallelism techniques are essential for efficiently training large language models (LLMs). Nevertheless, current automatic parallel planning frameworks often overlook the simultaneous consideration of node hetero…

Cloud Computing

Parallelization of a new embedded application for automatic meteor detection

2023-07-20 · Mathuran Kandeepan, Clara Ciocan, Adrien Cassagne, Lionel Lacassagne

This article presents the methods used to parallelize a new computer vision application. The system is able to automatically detect meteor from non-stabilized cameras and noisy video sequences. The application is designe…

Raspberry Pi 4