paper-with-me

홈 › Papers

Autonomous state-space segmentation for Deep-RL sparse reward scenarios

2025-04-04 · Gianluca Maselli, Vieri Giuliano Santucci

Dealing with environments with sparse rewards has always been crucial for systems developed to operate in autonomous open-ended learning settings. Intrinsic Motivations could be an effective way to help Deep Reinforcement Learning algorithms learn in such scenarios. In fact, intrinsic reward signals, such as novelty or curiosity, are generally adopted to improve exploration when extrinsic rewards are delayed or absent. Building on previous works, we tackle the problem of learning policies in the presence of sparse rewards by proposing a two-level architecture that alternates an ''intrinsically driven'' phase of exploration and autonomous sub-goal generation, to a phase of sparse reward, goal-directed policy learning. The idea is to build several small networks, each one specialized on a particular sub-path, and use them as starting points for future exploration without the need to further explore from scratch previously learnt paths. Two versions of the system have been trained and tested in the Gym SuperMarioBros environment without considering any additional extrinsic reward. The results show the validity of our approach and the importance of autonomously segment the environment to generate an efficient path towards the final goal.

📄 PDF Abstract BibTeX arXiv:2504.03420

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement Learning

Similar Papers 제목 키워드 기반

Discovering and Exploiting Sparse Rewards in a Learned Behavior Space

2021-11-02 · Giuseppe Paolo, Miranda Coninx, Alban Laflaquière, Stephane Doncieux

Learning optimal policies in sparse rewards settings is difficult as the learning agent has little to no feedback on the quality of its actions. In these situations, a good strategy is to focus on exploration, hopefully …

Efficient Exploration

Learning in Sparse Rewards settings through Quality-Diversity algorithms

2022-03-02 · Giuseppe Paolo

In the Reinforcement Learning (RL) framework, the learning is guided through a reward signal. This means that in situations of sparse rewards the agent has to focus on exploration, in order to discover which action, or s…

DiversityReinforcement Learning (RL)

Driving Beyond Privilege: Distilling Dense-Reward Knowledge into Sparse-Reward Policies

2025-12-03 · Feeza Khan Khanzada, Jaerock Kwon arxiv

We study how to exploit dense simulator-defined rewards in vision-based autonomous driving without inheriting their misalignment with deployment metrics. In realistic simulators such as CARLA, privileged state (e.g., lan…

Reinforcement LearningAutonomous Driving

Beyond Rewards in Reinforcement Learning for Cyber Defence

2026-02-04 · Elizabeth Bates, Chris Hicks, Vasilios Mavroudis arxiv

Recent years have seen an explosion of interest in autonomous cyber defence agents trained to defend computer networks using deep reinforcement learning. These agents are typically trained in cyber gym environments using…

Reinforcement Learning

LEST: Large-scale LiDAR Semantic Segmentation with Transformer

2023-07-14 · Chuanyu Luo, Nuo Cheng, Sikun Ma, Han Li 외

Large-scale LiDAR-based point cloud semantic segmentation is a critical task in autonomous driving perception. Almost all of the previous state-of-the-art LiDAR semantic segmentation methods are variants of sparse 3D con…

Autonomous DrivingLIDAR Semantic SegmentationSegmentationSemantic Segmentation