paper-with-me

홈 › Papers

Generalizing Off-Policy Evaluation From a Causal Perspective For Sequential Decision-Making

2022-01-20 · Sonali Parbhoo, Shalmali Joshi, Finale Doshi-Velez

Assessing the effects of a policy based on observational data from a different policy is a common problem across several high-stake decision-making domains, and several off-policy evaluation (OPE) techniques have been proposed. However, these methods largely formulate OPE as a problem disassociated from the process used to generate the data (i.e. structural assumptions in the form of a causal graph). We argue that explicitly highlighting this association has important implications on our understanding of the fundamental limits of OPE. First, this implies that current formulation of OPE corresponds to a narrow set of tasks, i.e. a specific causal estimand which is focused on prospective evaluation of policies over populations or sub-populations. Second, we demonstrate how this association motivates natural desiderata to consider a general set of causal estimands, particularly extending the role of OPE for counterfactual off-policy evaluation at the level of individuals of the population. A precise description of the causal estimand highlights which OPE estimands are identifiable from observational data under the stated generative assumptions. For those OPE estimands that are not identifiable, the causal perspective further highlights where more experimental data is necessary, and highlights situations where human expertise can aid identification and estimation. Furthermore, many formalisms of OPE overlook the role of uncertainty entirely in the estimation process.We demonstrate how specifically characterising the causal estimand highlights the different sources of uncertainty and when human expertise can naturally manage this uncertainty. We discuss each of these aspects as actionable desiderata for future OPE research at scale and in-line with practical utility.

📄 PDF Abstract BibTeX arXiv:2201.08262

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualDecision MakingOff-policy evaluationSequential Decision Making

Similar Papers 제목 키워드 기반

A Survey of Methods, Challenges and Perspectives in Causality

2023-02-01 · Gaël Gendron, Michael Witbrock, Gillian Dobbie

Deep Learning models have shown success in a large variety of tasks by extracting correlation patterns from high-dimensional data but still struggle when generalizing out of their initial distribution. As causal engines …

Deep LearningSurvey

Fed-CausalDiff: Decoupled Synchronization for Federated Do-Simulation and Policy Evaluation

2026-06-21 · Pengfei Li, Mohammad Khalil arxiv

While federated learning enables collaborative modelling on decentralised data, standard methods merely fit historical observations. This purely observational approach is fundamentally insufficient for interventional inf…

Federated Learning

Amortized Active Causal Induction with Deep Reinforcement Learning

2024-05-26 · Yashas Annadani, Panagiotis Tigas, Stefan Bauer, Adam Foster

We present Causal Amortized Active Structure Learning (CAASL), an active intervention design policy that can select interventions that are adaptive, real-time and that does not require access to the likelihood. This poli…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningZero-shot Generalization

Counterfactual Evaluation of Slate Recommendations with Sequential Reward Interactions

2020-07-25 · James McInerney, Brian Brost, Praveen Chandar, Rishabh Mehrotra 외

Users of music streaming, video streaming, news recommendation, and e-commerce services often engage with content in a sequential manner. Providing and evaluating good sequences of recommendations is therefore a central …

counterfactualNews RecommendationOff-policy evaluationRecommendation Systems

Causality-Aware Transformer Networks for Robotic Navigation

2024-09-04 · Ruoyu Wang, Yao Liu, Yuanjiang Cao, Lina Yao

Current research in Visual Navigation reveals opportunities for improvement. First, the direct adoption of RNNs and Transformers often overlooks the specific differences between Embodied AI and traditional sequential dat…

Visual Navigation