paper-with-me

홈 › Papers

Efficient and Encrypted Inference using Binarized Neural Networks within In-Memory Computing Architectures

2025-10-27 · Gokulnath Rajendran, Suman Deb, Anupam Chattopadhyay arxiv

Binarized Neural Networks (BNNs) are a class of deep neural networks designed to utilize minimal computational resources, which drives their popularity across various applications. Recent studies highlight the potential of mapping BNN model parameters onto emerging non-volatile memory technologies, specifically using crossbar architectures, resulting in improved inference performance compared to traditional CMOS implementations. However, the common practice of protecting model parameters from theft attacks by storing them in an encrypted format and decrypting them at runtime introduces significant computational overhead, thus undermining the core principles of in-memory computing, which aim to integrate computation and storage. This paper presents a robust strategy for protecting BNN model parameters, particularly within in-memory computing frameworks. Our method utilizes a secret key derived from a physical unclonable function to transform model parameters prior to storage in the crossbar. Subsequently, the inference operations are performed on the encrypted weights, achieving a very special case of Fully Homomorphic Encryption (FHE) with minimal runtime overhead. Our analysis reveals that inference conducted without the secret key results in drastically diminished performance, with accuracy falling below 15%. These results validate the effectiveness of our protection strategy in securing BNNs within in-memory computing architectures while preserving computational efficiency.

📄 PDF Abstract BibTeX arXiv:2510.23034

Code (0)

등록된 구현이 없습니다.

Tasks

Computational Efficiency

Similar Papers 제목 키워드 기반

BILLNET: A Binarized Conv3D-LSTM Network with Logic-gated residual architecture for hardware-efficient video inference

2025-01-24 · Van Thien Nguyen, William Guicquero, Gilles Sicard

Long Short-Term Memory (LSTM) and 3D convolution (Conv3D) show impressive results for many video-based applications but require large memory and intensive computing. Motivated by recent works on hardware-algorithmic co-d…

Enabling Homomorphically Encrypted Inference for Large DNN Models

2021-03-30 · Guillermo Lloret-Talavera, Marc Jorda, Harald Servat, Fabian Boemer 외

The proliferation of machine learning services in the last few years has raised data privacy concerns. Homomorphic encryption (HE) enables inference using encrypted data but it incurs 100x-10,000x memory and runtime over…

MOBIUS: Model-Oblivious Binarized Neural Networks

2018-11-29 · Hiromasa Kitai, Jason Paul Cruz, Naoto Yanai, Naohisa Nishida 외

A privacy-preserving framework in which a computational resource provider receives encrypted data from a client and returns prediction results without decrypting the data, i.e., oblivious neural network or encrypted pred…

BIG-bench Machine LearningmodelPredictionPrivacy Preserving

An In-Memory Analog Computing Co-Processor for Energy-Efficient CNN Inference on Mobile Devices

2021-05-24 · Mohammed Elbtity, Abhishek Singh, Brendan Reidy, Xiaochen Guo 외

In this paper, we develop an in-memory analog computing (IMAC) architecture realizing both synaptic behavior and activation functions within non-volatile memory arrays. Spin-orbit torque magnetoresistive random-access me…

CPU

Embedded Binarized Neural Networks

2017-09-06 · Bradley McDanel, Surat Teerapittayanon, H. T. Kung

We study embedded Binarized Neural Networks (eBNNs) with the aim of allowing current binarized neural networks (BNNs) in the literature to perform feedforward inference efficiently on small embedded devices. We focus on …