paper-with-me

Papers

TREX: Trajectory Explanations for Multi-Objective Reinforcement Learning

2026-03-23 · Dilina Rajapakse, Juan C. Rosero, Ivana Dusparic arxiv

Reinforcement Learning (RL) has demonstrated its ability to solve complex decision-making problems in a variety of domains, by optimizing reward signals obtained through interaction with an environment. However, many real-world scenarios involve multiple, potentially conflicting objectives that cannot be easily represented by a single scalar reward. Multi-Objective Reinforcement Learning (MORL) addresses this limitation by enabling agents to optimize several objectives simultaneously, explicitly reasoning about trade-offs between them. However, the ``black box" nature of the RL models makes the decision process behind chosen objective trade-offs unclear. Current Explainable Reinforcement Learning (XRL) methods are typically designed for single scalar rewards and do not account for explanations with respect to distinct objectives or user preferences. To address this gap, in this paper we propose TREX, a Trajectory based Explainability framework to explain Multi-objective Reinforcement Learning policies, based on trajectory attribution. TREX generates trajectories directly from the learned expert policy, across different user preferences and clusters them into semantically meaningful temporal segments. We quantify the influence of these behavioural segments on the Pareto trade-off by training complementary policies that exclude specific clusters, measuring the resulting relative deviation on the observed rewards and actions compared to the original expert policy. Experiments on multi-objective MuJoCo environments - HalfCheetah, Ant and Swimmer, demonstrate the framework's ability to isolate and quantify the specific behavioural patterns.

📄 PDF Abstract BibTeX arXiv:2603.21988

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

TREX: Tree-Ensemble Representer-Point Explanations

2020-09-11 · Jonathan Brophy, Daniel Lowd

How can we identify the training examples that contribute most to the prediction of a tree ensemble? In this paper, we introduce TREX, an explanation system that provides instance-attribution explanations for tree ensemb…

BELLATREX: Building Explanations through a LocaLly AccuraTe Rule EXtractor

2022-03-29 · Klest Dedja, Felipe Kenji Nakano, Konstantinos Pliakos, Celine Vens

Tree-ensemble algorithms, such as random forest, are effective machine learning methods popular for their flexibility, high performance, and robustness to overfitting. However, since multiple learners are combined, they …

Binary ClassificationMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION

Prediction of Rectal Cancer Regrowth from Longitudinal Endoscopy

2026-05-13 · Jorge Tapias Gomez, Despoina Kanata, Aneesh Rangnekar, Christina Lee 외 arxiv

Clinical trial studies indicate benefit of watch-and-wait (WW) surveillance for patients with rectal cancer showing a complete or near clinical response (CR) directly after treatment (restaging). However, there are no ob…

Topology Adaptive Graph Estimation in High Dimensions

2014-10-27 · Johannes Lederer, Christian Müller

We introduce Graphical TREX (GTREX), a novel method for graph estimation in high-dimensional Gaussian graphical models. By conducting neighborhood selection with TREX, GTREX avoids tuning parameters and is adaptive to th…

Vocal Bursts Intensity Prediction

Non-convex Global Minimization and False Discovery Rate Control for the TREX

2016-04-22 · Jacob Bien, Irina Gaynanova, Johannes Lederer, Christian Müller

The TREX is a recently introduced method for performing sparse high-dimensional regression. Despite its statistical promise as an alternative to the lasso, square-root lasso, and scaled lasso, the TREX is computationally…