paper-with-me

홈 › Papers

A Simpler Alternative to Variational Regularized Counterfactual Risk Minimization

2024-09-15 · Hua Chang Bakker, Shashank Gupta, Harrie Oosterhuis

Variance regularized counterfactual risk minimization (VRCRM) has been proposed as an alternative off-policy learning (OPL) method. VRCRM method uses a lower-bound on the $f$-divergence between the logging policy and the target policy as regularization during learning and was shown to improve performance over existing OPL alternatives on multi-label classification tasks. In this work, we revisit the original experimental setting of VRCRM and propose to minimize the $f$-divergence directly, instead of optimizing for the lower bound using a $f$-GAN approach. Surprisingly, we were unable to reproduce the results reported in the original setting. In response, we propose a novel simpler alternative to f-divergence optimization by minimizing a direct approximation of f-divergence directly, instead of a $f$-GAN based lower bound. Experiments showed that minimizing the divergence using $f$-GANs did not work as expected, whereas our proposed novel simpler alternative works better empirically.

📄 PDF Abstract BibTeX arXiv:2409.09819

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION

Similar Papers 제목 키워드 기반

Variance Regularized Counterfactual Risk Minimization via Variational Divergence Minimization

2018-01-01 · ICLR 2018 1 · Hang Wu

Off-policy learning, the task of evaluating and improving policies using historic data collected from a logging policy, is important because on-policy evaluation is usually expensive and has adverse impacts. One of the m…

counterfactual

Adversary-Free Counterfactual Prediction via Information-Regularized Representations

2025-10-17 · Shiqin Tang, Rong Feng, Shuxin Zhuang, Youzhi Zhang 외 arxiv

We study counterfactual prediction under assignment bias and propose a mathematically grounded, information-theoretic approach that removes treatment-covariate dependence without adversarial training. Starting from a bou…

From Variational to Deterministic Autoencoders

2019-03-29 · ICLR 2020 1 · Partha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael Black 외

Variational Autoencoders (VAEs) provide a theoretically-backed and popular framework for deep generative models. However, learning a VAE from data poses still unanswered theoretical questions and considerable practical c…

DecoderDensity Estimation

A Hybrid Enumeration Framework for Optimal Counterfactual Generation in Post-Acute COVID-19 Heart Failure

2025-10-21 · Jingya Cheng, Alaleh Azhir, Jiazi Tian, Hossein Estiri arxiv

Counterfactual inference provides a mathematical framework for reasoning about hypothetical outcomes under alternative interventions, bridging causal reasoning and predictive modeling. We present a counterfactual inferen…

Bayesian Counterfactual Risk Minimization

2018-06-29 · Ben London, Ted Sandler

We present a Bayesian view of counterfactual risk minimization (CRM) for offline learning from logged bandit feedback. Using PAC-Bayesian analysis, we derive a new generalization bound for the truncated inverse propensit…

counterfactual