paper-with-me

홈 › Papers

Residual Learning and Context Encoding for Adaptive Offline-to-Online Reinforcement Learning

2024-06-12 · Mohammadreza Nakhaei, Aidan Scannell, Joni Pajarinen

Offline reinforcement learning (RL) allows learning sequential behavior from fixed datasets. Since offline datasets do not cover all possible situations, many methods collect additional data during online fine-tuning to improve performance. In general, these methods assume that the transition dynamics remain the same during both the offline and online phases of training. However, in many real-world applications, such as outdoor construction and navigation over rough terrain, it is common for the transition dynamics to vary between the offline and online phases. Moreover, the dynamics may vary during the online fine-tuning. To address this problem of changing dynamics from offline to online RL we propose a residual learning approach that infers dynamics changes to correct the outputs of the offline solution. At the online fine-tuning phase, we train a context encoder to learn a representation that is consistent inside the current online learning environment while being able to predict dynamic transitions. Experiments in D4RL MuJoCo environments, modified to support dynamics' changes upon environment resets, show that our approach can adapt to these dynamic changes and generalize to unseen perturbations in a sample-efficient way, whilst comparison methods cannot.

📄 PDF Abstract BibTeX arXiv:2406.08238

Code (1)

mohammadrezanakhaei/relce 공식 구현 pytorch

Tasks

D4RLMuJoCoReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Online Residual Learning from Offline Experts for Pedestrian Tracking

2024-09-06 · Anastasios Vlachos, Anastasios Tsiamis, Aren Karapetyan, Efe C. Balta 외

In this paper, we consider the problem of predicting unknown targets from data. We propose Online Residual Learning (ORL), a method that combines online adaptation with offline-trained predictions. At a lower level, we e…

Pedestrian Trajectory PredictionPredictionTrajectory Prediction

ARMFlow: AutoRegressive MeanFlow for Online 3D Human Reaction Generation

2025-12-18 · Zichen Geng, Zeeshan Hayder, Wei Liu, Hesheng Wang 외 arxiv

3D human reaction generation faces three main challenges:(1) high motion fidelity, (2) real-time inference, and (3) autoregressive adaptability for online scenarios. Existing methods fail to meet all three simultaneously…

RESAR-BEV: An Explainable Progressive Residual Autoregressive Approach for Camera-Radar Fusion in BEV Segmentation

2025-05-10 · Zhiwen Zeng, Yunfei Yin, Zheng Yuan, Argho Dey 외

Bird's-Eye-View (BEV) semantic segmentation provides comprehensive environmental perception for autonomous driving but suffers multi-modal misalignment and sensor noise. We propose RESAR-BEV, a progressive refinement fra…

Autonomous DrivingBEV SegmentationSemantic Segmentation

Debiased Offline Representation Learning for Fast Online Adaptation in Non-stationary Dynamics

2024-02-17 · Xinyu Zhang, Wenjie Qiu, Yi-Chen Li, Lei Yuan 외

Developing policies that can adjust to non-stationary environments is essential for real-world reinforcement learning applications. However, learning such adaptable policies in offline settings, with only a limited set o…

MuJoCoRepresentation Learning

INS-Conv: Incremental Sparse Convolution for Online 3D Segmentation

2022-01-01 · CVPR 2022 1 · Leyao Liu, Tian Zheng, Yun-Jou Lin, Kai Ni 외

We propose INS-Conv, an INcremental Sparse Convolutional network which enables online accurate 3D semantic and instance segmentation. Benefiting from the incremental nature of RGB-D reconstruction, we only need to up…

CPUGPUInstance SegmentationRGB-D Reconstruction+2