paper-with-me

Papers

Causal policy ranking

2021-11-16 · Daniel McNamee, Hana Chockler

Policies trained via reinforcement learning (RL) are often very complex even for simple tasks. In an episode with $n$ time steps, a policy will make $n$ decisions on actions to take, many of which may appear non-intuitive to the observer. Moreover, it is not clear which of these decisions directly contribute towards achieving the reward and how significant is their contribution. Given a trained policy, we propose a black-box method based on counterfactual reasoning that estimates the causal effect that these decisions have on reward attainment and ranks the decisions according to this estimate. In this preliminary work, we compare our measure against an alternative, non-causal, ranking procedure, highlight the benefits of causality-based policy ranking, and discuss potential future work integrating causal algorithms into the interpretation of RL agent policies.

📄 PDF Abstract BibTeX arXiv:2111.08415

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualCounterfactual ReasoningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Metalearners for Ranking Treatment Effects

2024-05-03 · Toon Vanderschueren, Wouter Verbeke, Felipe Moraes, Hugo Manuel Proença

Efficiently allocating treatments with a budget constraint constitutes an important challenge across various domains. In marketing, for example, the use of promotions to target potential customers and boost conversions i…

Causal InferenceLearning-To-RankMarketing

Unconfounded Propensity Estimation for Unbiased Ranking

2023-05-17 · Dan Luo, Lixin Zou, Qingyao Ai, Zhiyu Chen 외

The goal of unbiased learning to rank (ULTR) is to leverage implicit user feedback for optimizing learning-to-rank systems. Among existing solutions, automatic ULTR algorithms that jointly learn user bias models (i.e., p…

Learning-To-Rank

Causal Judge Evaluation: Calibrated Surrogate Metrics for LLM Systems

2025-12-11 · Eddie Landesberg, Manjari Narayan arxiv

Measuring long-run LLM outcomes (user satisfaction, expert judgment, downstream KPIs) is expensive. Teams default to cheap LLM judges, but uncalibrated proxies can invert rankings entirely. Causal Judge Evaluation (CJE) …

Causal intersectionality for fair ranking

2020-06-15 · Ke Yang, Joshua R. Loftus, Julia Stoyanovich

In this paper we propose a causal modeling approach to intersectional fairness, and a flexible, task-specific method for computing intersectionally fair rankings. Rankings are used in many contexts, ranging from Web sear…

Causal InferenceFairness

Beyond Simulation: Benchmarking World Models for Planning and Causality in Autonomous Driving

2025-08-03 · Hunter Schofield, Mohammed Elmahgiubi, Kasra Rezaee, Jinjun Shan arxiv

World models have become increasingly popular in acting as learned traffic simulators. Recent work has explored replacing traditional traffic simulators with world models for policy training. In this work, we explore the…

Autonomous Driving