Combining Deep Learning and Model-Based Methods for Robust Real-Time Semantic Landmark Detection
Compared to abstract features, significant objects, so-called landmarks, are a more natural means for vehicle localization and navigation, especially in challenging unstructured environments. The major challenge is to recognize landmarks in various lighting conditions and changing environment (growing vegetation) while only having few training samples available. We propose a new method which leverages Deep Learning as well as model-based methods to overcome the need of a large data set. Using RGB images and light detection and ranging (LiDAR) point clouds, our approach combines state-of-the-art classification results of Convolutional Neural Networks (CNN), with robust model-based methods by taking prior knowledge of previous time steps into account. Evaluations on a challenging real-wold scenario, with trees and bushes as landmarks, show promising results over pure learning-based state-of-the-art 3D detectors, while being significant faster.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Real-Time Shape Tracking of Facial Landmarks
Detection of facial landmarks and accurate tracking of their shape are essential in real-time virtual makeup applications, where users can see the makeups effect by moving their face in different directions. Typical face…
Semantic SegmentationVisual Semantic SLAM with Landmarks for Large-Scale Outdoor Environment
Semantic SLAM is an important field in autonomous driving and intelligent agents, which can enable robots to achieve high-level navigation tasks, obtain simple cognition or reasoning ability and achieve language-based hu…
Autonomous DrivingSemantic SegmentationSemantic SLAMSChME at SemEval-2020 Task 1: A Model Ensemble for Detecting Lexical Semantic Change
This paper describes SChME (Semantic Change Detection with Model Ensemble), a method usedin SemEval-2020 Task 1 on unsupervised detection of lexical semantic change. SChME usesa model ensemble combining signals of distri…
Change DetectionWord EmbeddingsCombining Data-driven and Model-driven Methods for Robust Facial Landmark Detection
Facial landmark detection is an important yet challenging task for real-world computer vision applications. This paper proposes an effective and robust approach for facial landmark detection by combining data- and model-…
Facial Landmark DetectionTowers of Babel: Combining Images, Language, and 3D Geometry for Learning Multimodal Vision
The abundance and richness of Internet photos of landmarks and cities has led to significant progress in 3D vision over the past two decades, including automated 3D reconstructions of the world's landmarks from tourist p…
3D geometryDescriptiveImage CaptioningMultimodal Reasoning