Stereo Vision-based Semantic 3D Object and Ego-motion Tracking for Autonomous Driving
We propose a stereo vision-based approach for tracking the camera ego-motion and 3D semantic objects in dynamic autonomous driving scenarios. Instead of directly regressing the 3D bounding box using end-to-end approaches, we propose to use the easy-to-labeled 2D detection and discrete viewpoint classification together with a light-weight semantic inference method to obtain rough 3D object measurements. Based on the object-aware-aided camera pose tracking which is robust in dynamic environments, in combination with our novel dynamic object bundle adjustment (BA) approach to fuse temporal sparse feature correspondences and the semantic 3D measurement model, we obtain 3D object pose, velocity and anchored dynamic point cloud estimation with instance accuracy and temporal consistency. The performance of our proposed method is demonstrated in diverse scenarios. Both the ego-motion estimation and object localization are compared with the state-of-of-the-art solutions.
Code (0)
등록된 구현이 없습니다.
Tasks
Autonomous DrivingMotion EstimationObjectObject LocalizationPose TrackingSimilar Papers 제목 키워드 기반
Stereo Vision for Unmanned Aerial VehicleDetection, Tracking, and Motion Control
An innovative method of detecting Unmanned Aerial Vehicles (UAVs) is presented. The goal of this study is to develop a robust setup for an autonomous multi-rotor hunter UAV, capable of visually detecting and tracking the…
Motion Planningobject-detectionObject DetectionReconstructing 3D Motion Trajectory of Large Swarm of Flying Objects
This paper addresses the problem of reconstructing the motion trajectories of the individuals in a large collection of flying objects using two temporally synchronized and geometrically calibrated cameras. The 3D traject…
Stereo MatchingReal-Time Model-Based Rigid Object Pose Estimation and Tracking Combining Dense and Sparse Visual Cues
We propose a novel model-based method for estimating and tracking the six-degrees-of-freedom (6DOF) pose of rigid objects of arbitrary shapes in real-time. By combining dense motion and stereo cues with sparse keypoint c…
Pose EstimationView-Invariant Localization using Semantic Objects in Changing Environments
This paper proposes a novel framework for real-time localization and egomotion tracking of a vehicle in a reference map. The core idea is to map the semantic objects observed by the vehicle and register them to their cor…
PositionEmpowering Dynamic Urban Navigation with Stereo and Mid-Level Vision
The success of foundation models in language and vision motivated research in fully end-to-end robot navigation foundation models (NFMs). NFMs directly map monocular visual input to control actions and ignore mid-level v…
Spatial ReasoningDepth EstimationRobot Navigation