paper-with-me

홈 › Papers

Active Perception with Initial-State Uncertainty: A Policy Gradient Method

2024-09-24 · Chongyang Shi, Shuo Han, Michael Dorothy, Jie Fu

This paper studies the synthesis of an active perception policy that maximizes the information leakage of the initial state in a stochastic system modeled as a hidden Markov model (HMM). Specifically, the emission function of the HMM is controllable with a set of perception or sensor query actions. Given the goal is to infer the initial state from partial observations in the HMM, we use Shannon conditional entropy as the planning objective and develop a novel policy gradient method with convergence guarantees. By leveraging a variant of observable operators in HMMs, we prove several important properties of the gradient of the conditional entropy with respect to the policy parameters, which allow efficient computation of the policy gradient and stable and fast convergence. We demonstrate the effectiveness of our solution by applying it to an inference problem in a stochastic grid world environment.

📄 PDF Abstract BibTeX arXiv:2409.16439

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Real-World Reinforcement Learning of Active Perception Behaviors

2025-12-01 · Edward S. Hu, Jie Wang, Xingfang Yuan, Fiona Luo 외 arxiv

A robot's instantaneous sensory observations do not always reveal task-relevant state information. Under such partial observability, optimal behavior typically involves explicitly acting to gain the missing information. …

Reinforcement Learning

ISC-POMDPs: Partially Observed Markov Decision Processes with Initial-State Dependent Costs

2025-03-06 · Timothy L. Molloy

We introduce a class of partially observed Markov decision processes (POMDPs) with costs that can depend on both the value and (future) uncertainty associated with the initial state. These Initial-State Cost POMDPs (ISC-…

Robot Navigation

Active Perception in Adversarial Scenarios using Maximum Entropy Deep Reinforcement Learning

2019-02-14 · Macheng Shen, Jonathan P. How

We pose an active perception problem where an autonomous agent actively interacts with a second agent with potentially adversarial behaviors. Given the uncertainty in the intent of the other agent, the objective is to co…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Learning Vision-Driven Reactive Soccer Skills for Humanoid Robots

2025-11-06 · Yushi Wang, Changsheng Luo, Penghui Chen, Jianran Liu 외 arxiv

Humanoid soccer poses a representative challenge for embodied intelligence, requiring robots to coordinate agile locomotion with unreliable visual perception in dynamic environments. However, existing systems typically r…

Reinforcement Learning

Tube Diffusion Policy: Reactive Visual-Tactile Policy Learning for Contact-rich Manipulation

2026-04-26 · Teng Xue, Alberto Rigo, Bingjian Huang, Jiayi Shen 외 arxiv

Contact-rich manipulation is central to many everyday human activities, requiring continuous adaptation to contact uncertainty and external disturbances through multi-modal perception, particularly vision and tactile fee…