paper-with-me

홈 › Papers

Bingham Policy Parameterization for 3D Rotations in Reinforcement Learning

2022-02-08 · Stephen James, Pieter Abbeel

We propose a new policy parameterization for representing 3D rotations during reinforcement learning. Today in the continuous control reinforcement learning literature, many stochastic policy parameterizations are Gaussian. We argue that universally applying a Gaussian policy parameterization is not always desirable for all environments. One such case in particular where this is true are tasks that involve predicting a 3D rotation output, either in isolation, or coupled with translation as part of a full 6D pose output. Our proposed Bingham Policy Parameterization (BPP) models the Bingham distribution and allows for better rotation (quaternion) prediction over a Gaussian policy parameterization in a range of reinforcement learning tasks. We evaluate BPP on the rotation Wahba problem task, as well as a set of vision-based next-best pose robot manipulation tasks from RLBench. We hope that this paper encourages more research into developing other policy parameterization that are more suited for particular environments, rather than always assuming Gaussian.

📄 PDF Abstract BibTeX arXiv:2202.03957

Code (1)

stepjam/BPP 공식 구현 pytorch

Tasks

continuous-controlContinuous Controlreinforcement-learningReinforcement LearningReinforcement Learning (RL)Robot Manipulation

Similar Papers 제목 키워드 기반

Deep Bingham Networks: Dealing with Uncertainty and Ambiguity in Pose Estimation

2020-12-20 · Haowen Deng, Mai Bui, Nassir Navab, Leonidas Guibas 외

In this work, we introduce Deep Bingham Networks (DBN), a generic framework that can naturally handle pose-related uncertainties and ambiguities arising in almost all real life applications concerning 3D data. While exis…

Camera RelocalizationPose Estimation

Probabilistic Rotation Representation With an Efficiently Computable Bingham Loss Function and Its Application to Pose Estimation

2022-03-09 · Hiroya Sato, Takuya Ikeda, Koichi Nishiwaki

In recent years, a deep learning framework has been widely used for object pose estimation. While quaternion is a common choice for rotation representation of 6D pose, it cannot represent an uncertainty of the observatio…

Pose Estimation

L2Calib: $SE(3)$-Manifold Reinforcement Learning for Robust Extrinsic Calibration with Degenerate Motion Resilience

2025-08-08 · Baorun Li, Chengrui Zhu, Siyi Du, Bingran Chen 외 arxiv

Extrinsic calibration is essential for multi-sensor fusion, existing methods rely on structured targets or fully-excited data, limiting real-world applicability. Online calibration further suffers from weak excitation, l…

Reinforcement Learning

A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence

2023-01-30 · NeurIPS 2023 11 · Carlo Alfano, Rui Yuan, Patrick Rebeschini

Modern policy optimization methods in reinforcement learning, such as TRPO and PPO, owe their success to the use of parameterized policies. However, while theoretical guarantees have been established for this class of al…

Bingham Procrustean Alignment for Object Detection in Clutter

2013-04-27 · Jared Glover, Sanja Popovic

A new system for object detection in cluttered RGB-D images is presented. Our main contribution is a new method called Bingham Procrustean Alignment (BPA) to align models with the scene. BPA uses point correspondences be…

Objectobject-detectionObject DetectionPosition