paper-with-me

Papers

PI-QT-Opt: Predictive Information Improves Multi-Task Robotic Reinforcement Learning at Scale

2022-10-15 · Kuang-Huei Lee, Ted Xiao, Adrian Li, Paul Wohlhart, Ian Fischer, Yao Lu

The predictive information, the mutual information between the past and future, has been shown to be a useful representation learning auxiliary loss for training reinforcement learning agents, as the ability to model what will happen next is critical to success on many control tasks. While existing studies are largely restricted to training specialist agents on single-task settings in simulation, in this work, we study modeling the predictive information for robotic agents and its importance for general-purpose agents that are trained to master a large repertoire of diverse skills from large amounts of data. Specifically, we introduce Predictive Information QT-Opt (PI-QT-Opt), a QT-Opt agent augmented with an auxiliary loss that learns representations of the predictive information to solve up to 297 vision-based robot manipulation tasks in simulation and the real world with a single set of parameters. We demonstrate that modeling the predictive information significantly improves success rates on the training tasks and leads to better zero-shot transfer to unseen novel tasks. Finally, we evaluate PI-QT-Opt on real robots, achieving substantial and consistent improvement over QT-Opt in multiple experimental settings of varying environments, skills, and multi-task configurations.

📄 PDF Abstract BibTeX arXiv:2210.08217

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning (RL)Representation LearningRobot Manipulation

Similar Papers 제목 키워드 기반

Quasi-Periodic Gaussian Process Predictive Iterative Learning Control

2026-02-20 · Unnati Nigam, Radhendushka Srivastava, Faezeh Marzbanrad, Michael Burke arxiv

Repetitive motion tasks are common in robotics, but performance can degrade over time due to environmental changes and robot wear and tear. Iterative learning control (ILC) improves performance by using information from …

Gaussian Processes

Future Predictive Success-or-Failure Classification for Long-Horizon Robotic Tasks

2024-04-04 · Naoya Sogi, Hiroyuki Oyama, Takashi Shibata, Makoto Terao

Automating long-horizon tasks with a robotic arm has been a central research topic in robotics. Optimization-based action planning is an efficient approach for creating an action plan to complete a given task. Constructi…

ClassificationFuture prediction

Encoding Longer-term Contextual Multi-modal Information in a Predictive Coding Model

2018-04-17 · Junpei Zhong, Tetsuya OGATA, Angelo Cangelosi

Studies suggest that within the hierarchical architecture, the topological higher level possibly represents a conscious category of the current sensory events with slower changing activities. They attempt to predict the …

Path planning for unmanned surface vehicle based on predictive artificial potential field. International Journal of Advanced Robotic Systems

2026-02-22 · Jia Song, Ce Hao, Jiangcheng Su arxiv

Path planning for high-speed unmanned surface vehicles requires more complex solutions to reduce sailing time and save energy. This article proposes a new predictive artificial potential field that incorporates time info…

Sample Efficient Robot Learning in Supervised Effect Prediction Tasks

2024-12-03 · Mehmet Arda Eren, Erhan Oztop

In self-supervised robotic learning, agents acquire data through active interaction with their environment, incurring costs such as energy use, human oversight, and experimental time. To mitigate these, sample-efficient …

Active LearningDiversityEfficient ExplorationPrediction+1