Scene-level Pose Estimation for Multiple Instances of Densely Packed Objects
This paper introduces key machine learning operations that allow the realization of robust, joint 6D pose estimation of multiple instances of objects either densely packed or in unstructured piles from RGB-D data. The first objective is to learn semantic and instance-boundary detectors without manual labeling. An adversarial training framework in conjunction with physics-based simulation is used to achieve detectors that behave similarly in synthetic and real data. Given the stochastic output of such detectors, candidates for object poses are sampled. The second objective is to automatically learn a single score for each pose candidate that represents its quality in terms of explaining the entire scene via a gradient boosted tree. The proposed method uses features derived from surface and boundary alignment between the observed scene and the object model placed at hypothesized poses. Scene-level, multi-instance pose estimation is then achieved by an integer linear programming process that selects hypotheses that maximize the sum of the learned individual scores, while respecting constraints, such as avoiding collisions. To evaluate this method, a dataset of densely packed objects with challenging setups for state-of-the-art approaches is collected. Experiments on this dataset and a public one show that the method significantly outperforms alternatives in terms of 6D pose accuracy while trained only with synthetic datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
6D Pose EstimationPose EstimationSimilar Papers 제목 키워드 기반
Instance Influence Estimation for Hyperspectral Target Signature Characterization using Extended Functions of Multiple Instances
The Extended Functions of Multiple Instances (eFUMI) algorithm is a generalization of Multiple Instance Learning (MIL). In eFUMI, only bag level (i.e. set level) labels are needed to estimate target signatures from mixed…
Multiple Instance LearningPhysics-based Scene-level Reasoning for Object Pose Estimation in Clutter
This paper focuses on vision-based pose estimation for multiple rigid objects placed in clutter, especially in cases involving occlusions and objects resting on each other. Progress has been achieved recently in object r…
object-detectionObject DetectionObject RecognitionPose Estimation+1Normalized Object Coordinate Space for Category-Level 6D Object Pose and Size Estimation
The goal of this paper is to estimate the 6D pose and dimensions of unseen object instances in an RGB-D image. Contrary to "instance-level" 6D pose estimation tasks, our problem assumes that no exact object CAD models ar…
6D Pose Estimation6D Pose Estimation using RGBMixed RealityObject+1Review on 6D Object Pose Estimation with the focus on Indoor Scene Understanding
6D object pose estimation problem has been extensively studied in the field of Computer Vision and Robotics. It has wide range of applications such as robot manipulation, augmented reality, and 3D scene understanding. Wi…
6D Pose Estimation using RGBObjectPose EstimationRobot Manipulation+1Towards Scene Understanding with Detailed 3D Object Representations
Current approaches to semantic image and scene understanding typically employ rather simple object representations such as 2D or 3D bounding boxes. While such coarse models are robust and allow for reliable object detect…
3D Pose EstimationObjectobject-detectionObject Detection+2