paper-with-me

홈 › Papers

Active search and coverage using point-cloud reinforcement learning

2023-12-18 · Matthias Rosynski, Alexandru Pop, Lucian Busoniu

We consider a problem in which the trajectory of a mobile 3D sensor must be optimized so that certain objects are both found in the overall scene and covered by the point cloud, as fast as possible. This problem is called target search and coverage, and the paper provides an end-to-end deep reinforcement learning (RL) solution to solve it. The deep neural network combines four components: deep hierarchical feature learning occurs in the first stage, followed by multi-head transformers in the second, max-pooling and merging with bypassed information to preserve spatial relationships in the third, and a distributional dueling network in the last stage. To evaluate the method, a simulator is developed where cylinders must be found by a Kinect sensor. A network architecture study shows that deep hierarchical feature learning works for RL and that by using farthest point sampling (FPS) we can reduce the amount of points and achieve not only a reduction of the network size but also better results. We also show that multi-head attention for point-clouds helps to learn the agent faster but converges to the same outcome. Finally, we compare RL using the best network with a greedy baseline that maximizes immediate rewards and requires for that purpose an oracle that predicts the next observation. We decided RL achieves significantly better and more robust results than the greedy strategy.

📄 PDF Abstract BibTeX arXiv:2312.11410

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Double Q-learning Double Q-learning is an off-policy reinforcement learning algorithm that utilises double estimation to counteract overestimation problems with traditional Q-learning. The…
Dueling Network A Dueling Network is a type of Q-Network that has two streams to separately estimate (scalar) state-value and the advantages for each action. Both streams share a common…
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…

Similar Papers 제목 키워드 기반

IceCloudNet: Cirrus and mixed-phase cloud prediction from SEVIRI input learned from sparse supervision

2023-10-05 · Kai Jeggle, Mikolaj Czerkawski, Federico Serva, Bertrand Le Saux 외

Clouds containing ice particles play a crucial role in the climate system. Yet they remain a source of great uncertainty in climate models and future climate projections. In this work, we create a new observational const…

Continual learning on 3D point clouds with random compressed rehearsal

2022-05-16 · Maciej Zamorski, Michał Stypułkowski, Konrad Karanowski, Tomasz Trzciński 외

Contemporary deep neural networks offer state-of-the-art results when applied to visual reasoning, e.g., in the context of 3D point cloud data. Point clouds are important datatype for precise modeling of three-dimensiona…

Continual LearningVisual Reasoning

MARS: Multimodal Active Robotic Sensing for Articulated Characterization

2024-07-01 · Hongliang Zeng, Ping Zhang, Chengjiong Wu, Jiahua Wang 외

Precise perception of articulated objects is vital for empowering service robots. Recent studies mainly focus on point cloud, a single-modal approach, often neglecting vital texture and lighting details and assuming idea…

parameter estimation

PointPatchRL -- Masked Reconstruction Improves Reinforcement Learning on Point Clouds

2024-10-24 · Balázs Gyenes, Nikolai Franke, Philipp Becker, Gerhard Neumann

Perceiving the environment via cameras is crucial for Reinforcement Learning (RL) in robotics. While images are a convenient form of representation, they often complicate extracting important geometric details, especiall…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Representation Learning

Stein Variational Ergodic Surface Coverage with SE(3) Constraints

2026-03-10 · Jiayun Li, Yufeng Jin, Sangli Teng, Dejian Gong 외 arxiv

Surface manipulation tasks require robots to generate trajectories that comprehensively cover complex 3D surfaces while maintaining precise end-effector poses. Existing ergodic trajectory optimization (TO) methods demons…