EnforceNet: Monocular Camera Localization in Large Scale Indoor Sparse LiDAR Point Cloud
Pose estimation is a fundamental building block for robotic applications such as autonomous vehicles, UAV, and large scale augmented reality. It is also a prohibitive factor for those applications to be in mass production, since the state-of-the-art, centimeter-level pose estimation often requires long mapping procedures and expensive localization sensors, e.g. LiDAR and high precision GPS/IMU, etc. To overcome the cost barrier, we propose a neural network based solution to localize a consumer degree RGB camera within a prior sparse LiDAR map with comparable centimeter-level precision. We achieved it by introducing a novel network module, which we call resistor module, to enforce the network generalize better, predicts more accurately, and converge faster. Such results are benchmarked by several datasets we collected in the large scale indoor parking garage scenes. We plan to open both the data and the code for the community to join the effort to advance this field.
Code (0)
등록된 구현이 없습니다.
Tasks
Autonomous VehiclesCamera LocalizationPose EstimationSimilar Papers 제목 키워드 기반
Scale Drift Correction of Camera Geo-Localization using Geo-Tagged Images
Camera geo-localization from a monocular video is a fundamental task for video analysis and autonomous navigation. Although 3D reconstruction is a key technique to obtain camera poses, monocular 3D reconstruction in a la…
3D ReconstructionAutonomous Navigationgeo-localizationTranslationComparative Study of Vision-Based Metric Measurement for Large-Scale Planar Scenes
Vision-based metric distance and area measurement remains challenging in large-scale outdoor environments due to long-range sensing, camera zoom, and unstable imaging conditions. This work studies planar metric measureme…
Image StitchingImproved Real-Time Monocular SLAM Using Semantic Segmentation on Selective Frames
Monocular simultaneous localization and mapping (SLAM) is emerging in advanced driver assistance systems and autonomous driving, because a single camera is cheap and easy to install. Conventional monocular SLAM has two m…
Autonomous DrivingCPUGPUSegmentation+2BirdSLAM: Monocular Multibody SLAM in Bird's-Eye View
In this paper, we present BirdSLAM, a novel simultaneous localization and mapping (SLAM) system for the challenging scenario of autonomous driving platforms equipped with only a monocular camera. BirdSLAM tackles challen…
Autonomous DrivingMonocular ReconstructionObject LocalizationSimultaneous Localization and MappingUnsupervised Simultaneous Learning for Camera Re-Localization and Depth Estimation from Video
We present an unsupervised simultaneous learning framework for the task of monocular camera re-localization and depth estimation from unlabeled video sequences. Monocular camera re-localization refers to the task of esti…
Depth EstimationMonocular Depth Estimation