paper-with-me

Papers

Optical Computing for Deep Neural Network Acceleration: Foundations, Recent Developments, and Emerging Directions

2024-07-30 · Sudeep Pasricha

Emerging artificial intelligence applications across the domains of computer vision, natural language processing, graph processing, and sequence prediction increasingly rely on deep neural networks (DNNs). These DNNs require significant compute and memory resources for training and inference. Traditional computing platforms such as CPUs, GPUs, and TPUs are struggling to keep up with the demands of the increasingly complex and diverse DNNs. Optical computing represents an exciting new paradigm for light-speed acceleration of DNN workloads. In this article, we discuss the fundamentals and state-of-the-art developments in optical computing, with an emphasis on DNN acceleration. Various promising approaches are described for engineering optical devices, enhancing optical circuits, and designing architectures that can adapt optical computing to a variety of DNN workloads. Novel techniques for hardware/software co-design that can intelligently tune and map DNN models to improve performance and energy-efficiency on optical computing platforms across high performance and resource constrained embedded, edge, and IoT platforms are also discussed. Lastly, several open problems and future directions for research in this domain are highlighted.

📄 PDF Abstract BibTeX arXiv:2407.21184

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Unified Module for Accelerating STABLE-DIFFUSION: LCM-LORA

2024-03-24 · Ayush Thakur, Rashmi Vashisth

This paper presents a comprehensive study on the unified module for accelerating stable-diffusion processes, specifically focusing on the lcm-lora module. Stable-diffusion processes play a crucial role in various scienti…

Computational EfficiencyGPU

Cross-Layer Design for AI Acceleration with Non-Coherent Optical Computing

2023-03-22 · Febin Sunny, Mahdi Nikdast, Sudeep Pasricha

Emerging AI applications such as ChatGPT, graph convolutional networks, and other deep neural networks require massive computational resources for training and inference. Contemporary computing platforms such as CPUs, GP…

A Survey on Deep Learning Hardware Accelerators for Heterogeneous HPC Platforms

2023-06-27 · Cristina Silvano, Daniele Ielmini, Fabrizio Ferrandi, Leandro Fiorin 외

Recent trends in deep learning (DL) have made hardware accelerators essential for various high-performance computing (HPC) applications, including image classification, computer vision, and speech recognition. This surve…

Deep LearningGPUimage-classificationImage Classification+3

Training Large-Scale Optical Neural Networks with Two-Pass Forward Propagation

2024-08-15 · Amirreza Ahmadnejad, Somayyeh Koohi

This paper addresses the limitations in Optical Neural Networks (ONNs) related to training efficiency, nonlinear function implementation, and large input data processing. We introduce Two-Pass Forward Propagation, a nove…

Qibo: a framework for quantum simulation with hardware acceleration

2020-09-03 · Stavros Efthymiou, Sergi Ramos-Calderer, Carlos Bravo-Prieto, Adrián Pérez-Salinas 외

We present Qibo, a new open-source software for fast evaluation of quantum circuits and adiabatic evolution which takes full advantage of hardware accelerators. The growing interest in quantum computing and the recent de…

CPUGPU