paper-with-me

홈 › Papers

The Indoor-Training Effect: unexpected gains from distribution shifts in the transition function

2024-01-29 · Serena Bono, Spandan Madan, Ishaan Grover, Mao Yasueda, Cynthia Breazeal, Hanspeter Pfister, Gabriel Kreiman

Is it better to perform tennis training in a pristine indoor environment or a noisy outdoor one? To model this problem, here we investigate whether shifts in the transition probabilities between the training and testing environments in reinforcement learning problems can lead to better performance under certain conditions. We generate new Markov Decision Processes (MDPs) starting from a given MDP, by adding quantifiable, parametric noise into the transition function. We refer to this process as Noise Injection and the resulting environments as {\delta}-environments. This process allows us to create variations of the same environment with quantitative control over noise serving as a metric of distance between environments. Conventional wisdom suggests that training and testing on the same MDP should yield the best results. In stark contrast, we observe that agents can perform better when trained on the noise-free environment and tested on the noisy {\delta}-environments, compared to training and testing on the same {\delta}-environments. We confirm that this finding extends beyond noise variations: it is possible to showcase the same phenomenon in ATARI game variations including varying Ghost behaviour in PacMan, and Paddle behaviour in Pong. We demonstrate this intriguing behaviour across 60 different variations of ATARI games, including PacMan, Pong, and Breakout. We refer to this phenomenon as the Indoor-Training Effect. Code to reproduce our experiments and to implement Noise Injection can be found at https://bit.ly/3X6CTYk.

📄 PDF Abstract BibTeX arXiv:2401.15856

Code (0)

등록된 구현이 없습니다.

Tasks

Atari GamesReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Out of Distribution Detection via Domain-Informed Gaussian Process State Space Models

2023-09-13 · Alonso Marco, Elias Morley, Claire J. Tomlin

In order for robots to safely navigate in unseen scenarios using learning-based methods, it is important to accurately detect out-of-training-distribution (OoD) situations online. Recently, Gaussian process state-space m…

NavigateOut-of-Distribution DetectionState Space Models

DANDI: Diffusion as Normative Distribution for Deep Neural Network Input

2025-02-05 · SoMin Kim, Shin Yoo

Surprise Adequacy (SA) has been widely studied as a test adequacy metric that can effectively guide software engineers towards inputs that are more likely to reveal unexpected behaviour of Deep Neural Networks (DNNs). In…

DNN Testing

Decorrelated Clustering with Data Selection Bias

2020-06-29 · Xiao Wang, Shaohua Fan, Kun Kuang, Chuan Shi 외

Most of existing clustering algorithms are proposed without considering the selection bias in data. In many real applications, however, one cannot guarantee the data is unbiased. Selection bias might bring the unexpected…

ClusteringSelection bias

Navigating Beyond Instructions: Vision-and-Language Navigation in Obstructed Environments

2024-07-31 · Haodong Hong, Sen Wang, Zi Huang, Qi Wu 외

Real-world navigation often involves dealing with unexpected obstructions such as closed doors, moved objects, and unpredictable entities. However, mainstream Vision-and-Language Navigation (VLN) tasks typically assume i…

graph constructionNavigateVision and Language Navigation

On the value of distribution tail in the valuation of travel time variability

2022-07-13 · Zhaoqi Zang, Richard Batley, Xiangdong Xu, David Z. W. Wang

Extensive empirical studies show that the long distribution tail of travel time and the corresponding unexpected delay can have much more serious consequences than expected or moderate delay. However, the unexpected dela…