paper-with-me

홈 › Papers

STARC: A General Framework For Quantifying Differences Between Reward Functions

2023-09-26 · Joar Skalse, Lucy Farnik, Sumeet Ramesh Motwani, Erik Jenner, Adam Gleave, Alessandro Abate

In order to solve a task using reinforcement learning, it is necessary to first formalise the goal of that task as a reward function. However, for many real-world tasks, it is very difficult to manually specify a reward function that never incentivises undesirable behaviour. As a result, it is increasingly popular to use reward learning algorithms, which attempt to learn a reward function from data. However, the theoretical foundations of reward learning are not yet well-developed. In particular, it is typically not known when a given reward learning algorithm with high probability will learn a reward function that is safe to optimise. This means that reward learning algorithms generally must be evaluated empirically, which is expensive, and that their failure modes are difficult to anticipate in advance. One of the roadblocks to deriving better theoretical guarantees is the lack of good methods for quantifying the difference between reward functions. In this paper we provide a solution to this problem, in the form of a class of pseudometrics on the space of all reward functions that we call STARC (STAndardised Reward Comparison) metrics. We show that STARC metrics induce both an upper and a lower bound on worst-case regret, which implies that our metrics are tight, and that any metric with the same properties must be bilipschitz equivalent to ours. Moreover, we also identify a number of issues with reward metrics proposed by earlier works. Finally, we evaluate our metrics empirically, to demonstrate their practical efficacy. STARC metrics can be used to make both theoretical and empirical analysis of reward learning algorithms both easier and more principled.

📄 PDF Abstract BibTeX arXiv:2309.15257

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Complementarity, F-score, and NLP Evaluation

2016-05-01 · LREC 2016 5 · Leon Derczynski

This paper addresses the problem of quantifying the differences between entity extraction systems, where in general only a small proportion a document should be selected. Comparing overall accuracy is not very useful in …

Entity Extraction using GANInformation RetrievalRetrieval

Quantifying Social Biases in NLP: A Generalization and Empirical Comparison of Extrinsic Fairness Metrics

2021-06-28 · Paula Czarnowska, Yogarshi Vyas, Kashif Shah

Measuring bias is key for better understanding and addressing unfairness in NLP/ML models. This is often done via fairness metrics which quantify the differences in a model's behaviour across a range of demographic group…

Fairness

A Robust and Opponent-Aware League Training Method for StarCraft II

2023-09-21 · NeurIPS 2023 11

It is extremely difficult to train a superhuman Artificial Intelligence (AI) for games of similar size to StarCraft II. AlphaStar is the first AI that beat human professionals in the full game of StarCraft II, using a le…

pSTarC: Pseudo Source Guided Target Clustering for Fully Test-Time Adaptation

2023-09-02 · Manogna Sreenivas, Goirik Chakrabarty, Soma Biswas

Test Time Adaptation (TTA) is a pivotal concept in machine learning, enabling models to perform well in real-world scenarios, where test data distribution differs from training. In this work, we propose a novel approach …

ClusteringTest-time Adaptation

Elaboration and characterization of bioplastic films based on bitter cassava starch (Manihot esculenta) reinforced by chitosan extracted from crab (Shylla seratta) shells

2023-09-24 · Julie Tantely Mitantsoa, Pierre Hervé Ravelonandro, Fara Arimalala Andrianony, Rajaona Rafihavanana Andrianaivoravelona

Bioplastics are polymer plastics which are derived from renewable biomass resources. In this study, bioplastic films based on two different polysaccharides such as bitter cassava starch and chitosan extracted from crab s…