PixL2R: Guiding Reinforcement Learning Using Natural Language by Mapping Pixels to Rewards
Reinforcement learning (RL), particularly in sparse reward settings, often requires prohibitively large numbers of interactions with the environment, thereby limiting its applicability to complex problems. To address this, several prior approaches have used natural language to guide the agent's exploration. However, these approaches typically operate on structured representations of the environment, and/or assume some structure in the natural language commands. In this work, we propose a model that directly maps pixels to rewards, given a free-form natural language description of the task, which can then be used for policy learning. Our experiments on the Meta-World robot manipulation domain show that language-based rewards significantly improves the sample efficiency of policy learning, both in sparse and dense reward settings.
Code (1)
Tasks
reinforcement-learningReinforcement Learning (RL)Robot ManipulationSimilar Papers 제목 키워드 기반
PixLore: A Dataset-driven Approach to Rich Image Captioning
In the domain of vision-language integration, generating detailed image captions poses a significant challenge due to the lack of curated and rich datasets. This study introduces PixLore, a novel method that leverages Qu…
GPUImage CaptioningPixleepFlow: A Pixel-Based Lifelog Framework for Predicting Sleep Quality and Stress Level
The analysis of lifelogs can yield valuable insights into an individual's daily life, particularly with regard to their health and well-being. The accurate assessment of quality of life is necessitated by the use of dive…
Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)Sleep QualityPixLift: Accelerating Web Browsing via AI Upscaling
Accessing the internet in regions with expensive data plans and limited connectivity poses significant challenges, restricting information access and economic growth. Images, as a major contributor to webpage sizes, exac…
PIXLRelight: Controllable Relighting via Intrinsic Conditioning
We present PIXLRelight, a feed-forward approach for physically controllable single-image relighting. Existing methods either provide limited lighting control (e.g. through text or environment maps), accumulate errors whe…
3D ReconstructionImage RelightingIdentifying Human Edited Images using a CNN
Most non-professional photo manipulations are not made using propriety software like Adobe Photoshop, which is expensive and complicated to use for the average consumer selfie-taker or meme-maker. Instead, these individu…