paper-with-me

홈 › Papers

ALADIN: Accuracy-Latency-Aware Design-space Inference Analysis for Embedded AI Accelerators

2026-02-12 · T. Baldi, D. Casini, A. Biondi arxiv

The inference of deep neural networks (DNNs) on resource-constrained embedded systems introduces non-trivial trade-offs among model accuracy, computational latency, and hardware limitations, particularly when real-time constraints must be satisfied. This paper presents ALADIN, an accuracy-latency-aware design-space inference analysis framework for mixed-precision quantized neural networks (QNNs) targeting scratchpad-based AI accelerators. ALADIN enables the evaluation and analysis of inference bottlenecks and design trade-offs across accuracy, latency, and resource consumption without requiring deployment on the target platform, thereby significantly reducing development time and cost. The framework introduces a progressive refinement process that transforms a canonical QONNX model into platform-aware representations by integrating both platform-independent implementation details and hardware-specific characteristics. ALADIN is validated using a cycle-accurate simulator of a RISC-V based platform specialized for AI workloads, demonstrating its effectiveness as a tool for quantitative inference analysis and hardware-software co-design. Experimental results highlight how architectural decisions and mixed-precision quantization strategies impact accuracy, latency, and resource usage, and show that these effects can be precisely evaluated and compared using ALADIN, while also revealing subtle optimization tensions.

📄 PDF Abstract BibTeX arXiv:2603.08722

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Paladin: an annotation tool based on active and proactive learning

2021-04-01 · EACL 2021 2 · Minh-Quoc Nghiem, Paul Baylis, Sophia Ananiadou

In this paper, we present Paladin, an open-source web-based annotation tool for creating high-quality multi-label document-level datasets. By integrating active learning and proactive learning to the annotation task, Pal…

Active Learning

ALADIN: All Layer Adaptive Instance Normalization for Fine-grained Style Similarity

2021-03-17 · ICCV 2021 10 · Dan Ruta, Saeid Motiian, Baldo Faieta, Zhe Lin 외

We present ALADIN (All Layer AdaIN); a novel architecture for searching images based on the similarity of their artistic style. Representation learning is critical to visual search, where distance in the learned search e…

AllRepresentation Learning

Distributed Consensus Optimization with Consensus ALADIN

2025-03-21 · Xu Du, Jingzhe Wang

TThe paper proposes the Consensus Augmented Lagrange Alternating Direction Inexact Newton (Consensus ALADIN) algorithm, a novel approach for solving distributed consensus optimization problems (DC). Consensus ALADIN allo…

Computational Efficiency

ALADIN-$α$ -- An open-source MATLAB toolbox for distributed non-convex optimization

2020-06-02 · Alexander Engelmann, Yuning Jiang, Henrieke Benner, Ruchuan Ou 외

This paper introduces an open-source software for distributed and decentralized non-convex optimization named ALADIN-$\alpha$. ALADIN-$\alpha$ is a MATLAB implementation of tailored variants of the Augmented Lagrangian A…

Distributed Optimization

Convergence Theory of Flexible ALADIN for Distributed Optimization

2025-03-26 · Xu Du, Xiaohua Zhou, Shijie Zhu

The Augmented Lagrangian Alternating Direction Inexact Newton (ALADIN) method is a cutting-edge distributed optimization algorithm known for its superior numerical performance. It relies on each agent transmitting inform…

Distributed OptimizationFederated Learning