paper-with-me

Papers

VIBR: Learning View-Invariant Value Functions for Robust Visual Control

2023-06-14 · Tom Dupuis, Jaonary Rabarisoa, Quoc-Cuong Pham, David Filliat

End-to-end reinforcement learning on images showed significant progress in the recent years. Data-based approach leverage data augmentation and domain randomization while representation learning methods use auxiliary losses to learn task-relevant features. Yet, reinforcement still struggles in visually diverse environments full of distractions and spurious noise. In this work, we tackle the problem of robust visual control at its core and present VIBR (View-Invariant Bellman Residuals), a method that combines multi-view training and invariant prediction to reduce out-of-distribution (OOD) generalization gap for RL based visuomotor control. Our model-free approach improve baselines performances without the need of additional representation learning objectives and with limited additional computational cost. We show that VIBR outperforms existing methods on complex visuo-motor control environment with high visual perturbation. Our approach achieves state-of the-art results on the Distracting Control Suite benchmark, a challenging benchmark still not solved by current methods, where we evaluate the robustness to a number of visual perturbators, as well as OOD generalization and extrapolation capabilities.

📄 PDF Abstract BibTeX arXiv:2306.08537

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationRepresentation Learning

Similar Papers 제목 키워드 기반

Machine learning approach for vibronically renormalized electronic band structures

2024-09-03 · Niraj Aryal, Sheng Zhang, Weiguo Yin, Gia-Wei Chern

We present a machine learning (ML) method for efficient computation of vibrational thermal expectation values of physical properties from first principles. Our approach is based on the non-perturbative frozen phonon form…

Learning to Predict Structural Vibrations

2023-10-09 · Jan van Delden, Julius Schultz, Christopher Blech, Sabine C. Langer 외

In mechanical structures like airplanes, cars and houses, noise is generated and transmitted through vibrations. To take measures to reduce this noise, vibrations need to be simulated with expensive numerical computation…

Operator learningPDE Surrogate ModelingUncertainty Quantification

Policy Gradient Reinforcement Learning for Policy Represented by Fuzzy Rules: Application to Simulations of Speed Control of an Automobile

2020-09-04 · Seiji Ishihara, Harukazu Igarashi

A method of a fusion of fuzzy inference and policy gradient reinforcement learning has been proposed that directly learns, as maximizes the expected value of the reward per episode, parameters in a policy function repres…

Time Series Analysis

Learning Invariant Color Features for Person Re-Identification

2014-10-04 · Rahul Rama Varior, Gang Wang, Jiwen Lu

Matching people across multiple camera views known as person re-identification, is a challenging problem due to the change in visual appearance caused by varying lighting conditions. The perceived color of the subject ap…

Person Re-Identification

Complex-Valued GNNs for Distributed Basis-Invariant Control of Planar Systems

2026-04-03 · Samuel Honor, Mohamed Abdelnaby, Kevin Leahy arxiv

Graph neural networks (GNNs) are a well-regarded tool for learned control of networked dynamical systems due to their ability to be deployed in a distributed manner. However, current distributed GNN architectures assume …