The challenge of simultaneous object detection and pose estimation: a comparative study
Detecting objects and estimating their pose remains as one of the major challenges of the computer vision research community. There exists a compromise between localizing the objects and estimating their viewpoints. The detector ideally needs to be view-invariant, while the pose estimation process should be able to generalize towards the category-level. This work is an exploration of using deep learning models for solving both problems simultaneously. For doing so, we propose three novel deep learning architectures, which are able to perform a joint detection and pose estimation, where we gradually decouple the two tasks. We also investigate whether the pose estimation problem should be solved as a classification or regression problem, being this still an open question in the computer vision community. We detail a comparative analysis of all our solutions and the methods that currently define the state of the art for this problem. We use PASCAL3D+ and ObjectNet3D datasets to present the thorough experimental evaluation and main results. With the proposed models we achieve the state-of-the-art performance in both datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
object-detectionObject DetectionOpen-Ended Question AnsweringPose EstimationSimilar Papers 제목 키워드 기반
Simultaneous Multiple Object Detection and Pose Estimation using 3D Model Infusion with Monocular Vision
Multiple object detection and pose estimation are vital computer vision tasks. The latter relates to the former as a downstream problem in applications such as robotics and autonomous driving. However, due to the high co…
Autonomous DrivingObjectobject-detectionObject Detection+1A 3D Object Detection and Pose Estimation Pipeline Using RGB-D Images
3D object detection and pose estimation has been studied extensively in recent decades for its potential applications in robotics. However, there still remains challenges when we aim at detecting multiple objects while r…
3D Object DetectionClusteringObjectobject-detection+4DetFlowTrack: 3D Multi-object Tracking based on Simultaneous Optimization of Object Detection and Scene Flow Estimation
3D Multi-Object Tracking (MOT) is an important part of the unmanned vehicle perception module. Most methods optimize object detection and data association independently. These methods make the network structure complicat…
3D Multi-Object TrackingMulti-Object TrackingObjectobject-detection+3Self-Configurable Stabilized Real-Time Detection Learning for Autonomous Driving Applications
Guaranteeing real-time and accurate object detection simultaneously is paramount in autonomous driving environments. However, the existing object detection neural network systems are characterized by a tradeoff between c…
Autonomous DrivingObjectobject-detectionObject Detection+2Advanced Object Detection and Pose Estimation with Hybrid Task Cascade and High-Resolution Networks
In the field of computer vision, 6D object detection and pose estimation are critical for applications such as robotics, augmented reality, and autonomous driving. Traditional methods often struggle with achieving high a…
Autonomous DrivingObjectobject-detectionObject Detection+1