paper-with-me

홈 › Papers

Self-supervised Deep Reinforcement Learning with Generalized Computation Graphs for Robot Navigation

2017-09-29 · Gregory Kahn, Adam Villaflor, Bosen Ding, Pieter Abbeel, Sergey Levine

Enabling robots to autonomously navigate complex environments is essential for real-world deployment. Prior methods approach this problem by having the robot maintain an internal map of the world, and then use a localization and planning method to navigate through the internal map. However, these approaches often include a variety of assumptions, are computationally intensive, and do not learn from failures. In contrast, learning-based methods improve as the robot acts in the environment, but are difficult to deploy in the real-world due to their high sample complexity. To address the need to learn complex policies with few samples, we propose a generalized computation graph that subsumes value-based model-free methods and model-based methods, with specific instantiations interpolating between model-free and model-based. We then instantiate this graph to form a navigation model that learns from raw images and is sample efficient. Our simulated car experiments explore the design decisions of our navigation model, and show our approach outperforms single-step and $N$-step double Q-learning. We also evaluate our approach on a real-world RC car and show it can learn to navigate through a complex indoor environment with a few hours of fully autonomous, self-supervised training. Videos of the experiments and code can be found at github.com/gkahn13/gcg

📄 PDF Abstract BibTeX arXiv:1709.10489

Code (2)

gkahn13/gcg 공식 구현
abefetterman/hamstir-gym tf

Tasks

Deep Reinforcement LearningNavigateQ-Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Robot Navigation

Similar Papers 제목 키워드 기반

ProMerge: Prompt and Merge for Unsupervised Instance Segmentation

2024-09-27 · Dylan Li, Gyungin Shin

Unsupervised instance segmentation aims to segment distinct object instances in an image without relying on human-labeled data. This field has recently seen significant advancements, partly due to the strong local corres…

Instance SegmentationSemantic SegmentationUnsupervised Instance Segmentation

Self-supervised Graphs for Audio Representation Learning with Limited Labeled Data

2022-01-31 · Amir Shirian, Krishna Somandepalli, Tanaya Guha

Large scale databases with high-quality manual annotations are scarce in audio domain. We thus explore a self-supervised graph approach to learning audio representations from highly limited labelled data. Considering eac…

Emotion RecognitionEvent Detectiongraph constructionRepresentation Learning+1

Generalized Radiograph Representation Learning via Cross-supervision between Images and Free-text Radiology Reports

2021-11-04 · Hong-Yu Zhou, Xiaoyu Chen, Yinghao Zhang, Ruibang Luo 외

Pre-training lays the foundation for recent successes in radiograph analysis supported by deep learning. It learns transferable image representations by conducting large-scale fully-supervised or self-supervised learning…

Representation LearningSelf-Supervised LearningTransfer Learning

History-Aware Cross-Attention Reinforcement: Self-Supervised Multi Turn and Chain-of-Thought Fine-Tuning with vLLM

2025-06-08 · Andrew Kiruluta, Andreas Lemos, Priscilla Burity

We present CAGSR-vLLM-MTC, an extension of our Self-Supervised Cross-Attention-Guided Reinforcement (CAGSR) framework, now implemented on the high-performance vLLM runtime, to address both multi-turn dialogue and chain-o…

Learning Graph Representation by Aggregating Subgraphs via Mutual Information Maximization

2021-03-24 · Chenguang Wang, Ziwen Liu

In this paper, we introduce a self-supervised learning method to enhance the graph-level representations with the help of a set of subgraphs. For this purpose, we propose a universal framework to generate subgraphs in an…

AttributeContrastive LearningGraph Representation LearningRepresentation Learning+1