paper-with-me

Papers

QForce-RL: Quantized FPGA-Optimized Reinforcement Learning Compute Engine

2025-06-08 · Anushka Jha, Tanushree Dewangan, Mukul Lokhande, Santosh Kumar Vishvakarma

Reinforcement Learning (RL) has outperformed other counterparts in sequential decision-making and dynamic environment control. However, FPGA deployment is significantly resource-expensive, as associated with large number of computations in training agents with high-quality images and possess new challenges. In this work, we propose QForce-RL takes benefits of quantization to enhance throughput and reduce energy footprint with light-weight RL architecture, without significant performance degradation. QForce-RL takes advantages from E2HRL to reduce overall RL actions to learn desired policy and QuaRL for quantization based SIMD for hardware acceleration. We have also provided detailed analysis for different RL environments, with emphasis on model size, parameters, and accelerated compute ops. The architecture is scalable for resource-constrained devices and provide parametrized efficient deployment with flexibility in latency, throughput, power, and energy efficiency. The proposed QForce-RL provides performance enhancement up to 2.3x and better FPS - 2.6x compared to SoTA works.

📄 PDF Abstract BibTeX arXiv:2506.07046

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingQuantizationreinforcement-learningReinforcement LearningReinforcement Learning (RL)Sequential Decision Making

Similar Papers 제목 키워드 기반

Integer-only Quantized Transformers for Embedded FPGA-based Time-series Forecasting in AIoT

2024-07-06 · Tianheng Ling, Chao Qian, Gregor Schiele

This paper presents the design of a hardware accelerator for Transformers, optimized for on-device time-series forecasting in AIoT systems. It integrates integer-only quantization and Quantization-Aware Training with opt…

QuantizationTime SeriesTime Series Forecasting

End-to-end codesign of Hessian-aware quantized neural networks for FPGAs and ASICs

2023-04-13 · Javier Campos, Zhen Dong, Javier Duarte, Amir Gholami 외

We develop an end-to-end workflow for the training and implementation of co-designed neural networks (NNs) for efficient field-programmable gate array (FPGA) and application-specific integrated circuit (ASIC) hardware. O…

Quantization

Streaming Architecture for Large-Scale Quantized Neural Networks on an FPGA-Based Dataflow Platform

2017-07-31 · Chaim Baskin, Natan Liss, Evgenii Zheltonozhskii, Alex M. Bronshtein 외

Deep neural networks (DNNs) are used by different applications that are executed on a range of computer architectures, from IoT devices to supercomputers. The footprint of these networks is huge as well as their computat…

General Classification

AddNet: Deep Neural Networks Using FPGA-Optimized Multipliers

2019-11-19 · Julian Faraone, Martin Kumm, Martin Hardieck, Peter Zipf 외

Low-precision arithmetic operations to accelerate deep-learning applications on field-programmable gate arrays (FPGAs) have been studied extensively, because they offer the potential to save silicon area or increase thro…

Quantization

FlowPrecision: Advancing FPGA-Based Real-Time Fluid Flow Estimation with Linear Quantization

2024-03-04 · Tianheng Ling, Julian Hoever, Chao Qian, Gregor Schiele

In industrial and environmental monitoring, achieving real-time and precise fluid flow measurement remains a critical challenge. This study applies linear quantization in FPGA-based soft sensors for fluid flow estimation…

Quantization