paper-with-me

홈 › Papers

Generalized Simultaneous Perturbation-based Gradient Search with Reduced Estimator Bias

2022-12-20 · Soumen Pachal, Shalabh Bhatnagar, L. A. Prashanth

We present in this paper a family of generalized simultaneous perturbation-based gradient search (GSPGS) estimators that use noisy function measurements. The number of function measurements required by each estimator is guided by the desired level of accuracy. We first present in detail unbalanced generalized simultaneous perturbation stochastic approximation (GSPSA) estimators and later present the balanced versions (B-GSPSA) of these. We extend this idea further and present the generalized smoothed functional (GSF) and generalized random directions stochastic approximation (GRDSA) estimators, respectively, as well as their balanced variants. We show that estimators within any specified class requiring more number of function measurements result in lower estimator bias. We present a detailed analysis of both the asymptotic and non-asymptotic convergence of the resulting stochastic approximation schemes. We further present a series of experimental results with the various GSPGS estimators on the Rastrigin and quadratic function objectives. Our experiments are seen to validate our theoretical findings.

📄 PDF Abstract BibTeX arXiv:2212.10477

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Simultaneous Perturbation Algorithms for Batch Off-Policy Search

2014-03-18 · Raphael Fonteneau, L. A. Prashanth

We propose novel policy search algorithms in the context of off-policy, batch mode reinforcement learning (RL) with continuous state and action spaces. Given a batch collection of trajectories, we perform off-line policy…

Reinforcement LearningReinforcement Learning (RL)

Over-the-Air Federated Learning with Privacy Protection via Correlated Additive Perturbations

2022-10-05 · Jialing Liao, Zheng Chen, Erik G. Larsson

In this paper, we consider privacy aspects of wireless federated learning (FL) with Over-the-Air (OtA) transmission of gradient updates from multiple users/agents to an edge server. By exploiting the waveform superpositi…

Federated Learning

AdvCodeMix: Adversarial Attack on Code-Mixed Data

2021-10-30 · Sourya Dipta Das, Ayan Basak, Soumil Mandal, Dipankar Das

Research on adversarial attacks are becoming widely popular in the recent years. One of the unexplored areas where prior research is lacking is the effect of adversarial attacks on code-mixed data. Therefore, in the pres…

Adversarial AttackSentenceSentiment AnalysisSentiment Classification

Efficient Robustness Assessment via Adversarial Spatial-Temporal Focus on Videos

2023-01-03 · Wei Xingxing, Wang Songping, Yan Huanqian

Adversarial robustness assessment for video recognition models has raised concerns owing to their wide applications on safety-critical tasks. Compared with images, videos have much high dimension, which brings huge compu…

Action RecognitionAdversarial RobustnessMulti-agent Reinforcement LearningVideo Recognition

Gradient Estimation with Simultaneous Perturbation and Compressive Sensing

2015-11-27 · Vivek S. Borkar, Vikranth R. Dwaracherla, Neeraja Sahasrabudhe

This paper aims at achieving a "good" estimator for the gradient of a function on a high-dimensional space. Often such functions are not sensitive in all coordinates and the gradient of the function is almost sparse. We …

Compressive Sensing