Eliminating Ratio Bias for Gradient-based Simulated Parameter Estimation
This article addresses the challenge of parameter calibration in stochastic models where the likelihood function is not analytically available. We propose a gradient-based simulated parameter estimation framework, leveraging a multi-time scale algorithm that tackles the issue of ratio bias in both maximum likelihood estimation and posterior density estimation problems. Additionally, we introduce a nested simulation optimization structure, providing theoretical analyses including strong convergence, asymptotic normality, convergence rate, and budget allocation strategies for the proposed algorithm. The framework is further extended to neural network training, offering a novel perspective on stochastic approximation in machine learning. Numerical experiments show that our algorithm can improve the estimation accuracy and save computational costs.
Code (0)
등록된 구현이 없습니다.
Tasks
Density Estimationparameter estimationSimilar Papers 제목 키워드 기반
Safe, Scalable, and Accurate Bayes Posterior Sampling for Large-Data Generalized Linear Mixed Models
We consider the problem of scalable sampling algorithms to fit Bayesian generalized linear mixed models on large datasets. Stochastic gradient Langevin dynamics, coupled with smooth re-parameterizations of variance param…
Bayesian InferenceA New Stochastic Approximation Method for Gradient-based Simulated Parameter Estimation
This paper tackles the challenge of parameter calibration in stochastic models, particularly in scenarios where the likelihood function is unavailable in an analytical form. We introduce a gradient-based simulated parame…
Density Estimationparameter estimationVariational InferenceLearning-based MPC from Big Data Using Reinforcement Learning
This paper presents an approach for learning Model Predictive Control (MPC) schemes directly from data using Reinforcement Learning (RL) methods. The state-of-the-art learning methods use RL to improve the performance of…
Model Predictive Controlreinforcement-learningReinforcement LearningReinforcement Learning (RL)LEA: Label Enumeration Attack in Vertical Federated Learning
A typical Vertical Federated Learning (VFL) scenario involves several participants collaboratively training a machine learning model, where each party has different features for the same samples, with labels held exclusi…
Federated LearningEffects of Distributional Biases on Gradient-Based Causal Discovery in the Bivariate Categorical Case
Gradient-based causal discovery shows great potential for deducing causal structure from data in an efficient and scalable way. Those approaches however can be susceptible to distributional biases in the data they are tr…