paper-with-me

홈 › Papers

Learning to Navigate Using Mid-Level Visual Priors

2019-12-23 · Alexander Sax, Jeffrey O. Zhang, Bradley Emi, Amir Zamir, Silvio Savarese, Leonidas Guibas, Jitendra Malik

How much does having visual priors about the world (e.g. the fact that the world is 3D) assist in learning to perform downstream motor tasks (e.g. navigating a complex environment)? What are the consequences of not utilizing such visual priors in learning? We study these questions by integrating a generic perceptual skill set (a distance estimator, an edge detector, etc.) within a reinforcement learning framework (see Fig. 1). This skill set ("mid-level vision") provides the policy with a more processed state of the world compared to raw images. Our large-scale study demonstrates that using mid-level vision results in policies that learn faster, generalize better, and achieve higher final performance, when compared to learning from scratch and/or using state-of-the-art visual and non-visual representation learning methods. We show that conventional computer vision objectives are particularly effective in this regard and can be conveniently integrated into reinforcement learning frameworks. Finally, we found that no single visual representation was universally useful for all downstream tasks, hence we computationally derive a task-agnostic set of representations optimized to support arbitrary downstream tasks.

📄 PDF Abstract BibTeX arXiv:1912.11121

Code (1)

alexsax/midlevel-reps 공식 구현 pytorch

Tasks

Navigatereinforcement-learningReinforcement LearningReinforcement Learning (RL)Representation Learning

Similar Papers 제목 키워드 기반

Visual Semantic Navigation using Scene Priors

2018-10-15 · ICLR 2019 5 · Wei Yang, Xiaolong Wang, Ali Farhadi, Abhinav Gupta 외

How do humans navigate to target objects in novel scenes? Do we use the semantic/functional priors we have built over years to efficiently search and navigate? For example, to search for mugs, we search cabinets near the…

Deep Reinforcement LearningNavigateReinforcement Learning

Underexposed Image Correction via Hybrid Priors Navigated Deep Propagation

2019-07-17 · Risheng Liu, Long Ma, Yuxi Zhang, Xin Fan 외

Enhancing visual qualities for underexposed images is an extensively concerned task that plays important roles in various areas of multimedia and computer vision. Most existing methods often fail to generate high-quality…

Face DetectionSingle Image Haze Removal

Segment-Level Road Obstacle Detection Using Visual Foundation Model Priors and Likelihood Ratios

2024-12-07 · Youssef Shoeb, Nazir Nayal, Azarm Nowzard, Fatma Güney 외

Detecting road obstacles is essential for autonomous vehicles to navigate dynamic and complex traffic environments safely. Current road obstacle detection methods typically assign a score to each pixel and apply a thresh…

Autonomous VehiclesNavigate

Manipulate-to-Navigate: Reinforcement Learning with Visual Affordances and Manipulability Priors

2025-08-18 · Yuying Zhang, Joni Pajarinen arxiv

Mobile manipulation in dynamic environments is challenging due to movable obstacles blocking the robot's path. Traditional methods, which treat navigation and manipulation as separate tasks, often fail in such 'manipulat…

Reinforcement Learning

Learning Natural and Robust Hexapod Locomotion over Complex Terrains via Motion Priors based on Deep Reinforcement Learning

2025-11-05 · Xin Liu, Jinze Wu, Yinghui Li, Chenkun Qi 외 arxiv

Multi-legged robots offer enhanced stability to navigate complex terrains with their multiple legs interacting with the environment. However, how to effectively coordinate the multiple legs in a larger action exploration…

Reinforcement Learning