paper-with-me

Papers

Compute-in-Memory based Neural Network Accelerators for Safety-Critical Systems: Worst-Case Scenarios and Protections

2023-12-11 · Zheyu Yan, Xiaobo Sharon Hu, Yiyu Shi

Emerging non-volatile memory (NVM)-based Computing-in-Memory (CiM) architectures show substantial promise in accelerating deep neural networks (DNNs) due to their exceptional energy efficiency. However, NVM devices are prone to device variations. Consequently, the actual DNN weights mapped to NVM devices can differ considerably from their targeted values, inducing significant performance degradation. Many existing solutions aim to optimize average performance amidst device variations, which is a suitable strategy for general-purpose conditions. However, the worst-case performance that is crucial for safety-critical applications is largely overlooked in current research. In this study, we define the problem of pinpointing the worst-case performance of CiM DNN accelerators affected by device variations. Additionally, we introduce a strategy to identify a specific pattern of the device value deviations in the complex, high-dimensional value deviation space, responsible for this worst-case outcome. Our findings reveal that even subtle device variations can precipitate a dramatic decline in DNN accuracy, posing risks for CiM-based platforms in supporting safety-critical applications. Notably, we observe that prevailing techniques to bolster average DNN performance in CiM accelerators fall short in enhancing worst-case scenarios. In light of this issue, we propose a novel worst-case-aware training technique named A-TRICE that efficiently combines adversarial training and noise-injection training with right-censored Gaussian noise to improve the DNN accuracy in the worst-case scenarios. Our experimental results demonstrate that A-TRICE improves the worst-case accuracy under device variations by up to 33%.

📄 PDF Abstract BibTeX arXiv:2312.06137

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

When Small Variations Become Big Failures: Reliability Challenges in Compute-in-Memory Neural Accelerators

2026-03-03 · Yifan Qin, Jiahao Zheng, Zheyu Yan, Wujie Wen 외 arxiv

Compute-in-memory (CiM) architectures promise significant improvements in energy efficiency and throughput for deep neural network acceleration by alleviating the von Neumann bottleneck. However, their reliance on emergi…

Computing-In-Memory Neural Network Accelerators for Safety-Critical Systems: Can Small Device Variations Be Disastrous?

2022-07-15 · Zheyu Yan, Xiaobo Sharon Hu, Yiyu Shi

Computing-in-Memory (CiM) architectures based on emerging non-volatile memory (NVM) devices have demonstrated great potential for deep neural network (DNN) acceleration thanks to their high energy efficiency. However, NV…

ODMA: On-Demand Memory Allocation Strategy for LLM Serving on LPDDR-Class Accelerators

2025-12-10 · Guoqiang Zou, Wanyu Wang, Hao Zheng, Longxiang Yin 외 arxiv

Existing memory management techniques severely hinder efficient Large Language Model serving on accelerators constrained by poor random-access bandwidth.While static pre-allocation preserves memory contiguity,it incurs s…

Improving Realistic Worst-Case Performance of NVCiM DNN Accelerators through Training with Right-Censored Gaussian Noise

2023-07-29 · Zheyu Yan, Yifan Qin, Wujie Wen, Xiaobo Sharon Hu 외

Compute-in-Memory (CiM), built upon non-volatile memory (NVM) devices, is promising for accelerating deep neural networks (DNNs) owing to its in-situ data processing capability and superior energy efficiency. Unfortunate…

Self-Driving Cars

Efficient Error-Tolerant Quantized Neural Network Accelerators

2019-12-16 · Giulio Gambardella, Johannes Kappauf, Michaela Blott, Christoph Doehring 외

Neural Networks are currently one of the most widely deployed machine learning algorithms. In particular, Convolutional Neural Networks (CNNs), are gaining popularity and are evaluated for deployment in safety critical a…

QuantizationScheduling