paper-with-me

홈 › Papers

Real-Time Multi-Modal Semantic Fusion on Unmanned Aerial Vehicles with Label Propagation for Cross-Domain Adaptation

2022-10-18 · Simon Bultmann, Jan Quenzel, Sven Behnke

Unmanned aerial vehicles (UAVs) equipped with multiple complementary sensors have tremendous potential for fast autonomous or remote-controlled semantic scene analysis, e.g., for disaster examination. Here, we propose a UAV system for real-time semantic inference and fusion of multiple sensor modalities. Semantic segmentation of LiDAR scans and RGB images, as well as object detection on RGB and thermal images, run online onboard the UAV computer using lightweight CNN architectures and embedded inference accelerators. We follow a late fusion approach where semantic information from multiple sensor modalities augments 3D point clouds and image segmentation masks while also generating an allocentric semantic map. Label propagation on the semantic map allows for sensor-specific adaptation with cross-modality and cross-domain supervision. Our system provides augmented semantic images and point clouds with $\approx$ 9 Hz. We evaluate the integrated system in real-world experiments in an urban environment and at a disaster test site.

📄 PDF Abstract BibTeX arXiv:2210.09739

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationImage Segmentationobject-detectionObject DetectionSegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Real-Time Multi-Modal Semantic Fusion on Unmanned Aerial Vehicles

2021-08-14 · Simon Bultmann, Jan Quenzel, Sven Behnke

Unmanned aerial vehicles (UAVs) equipped with multiple complementary sensors have tremendous potential for fast autonomous or remote-controlled semantic scene analysis, e.g., for disaster examination. In this work, we pr…

Image Segmentationobject-detectionObject DetectionSegmentation+1

Missing Modality Robustness in Semi-Supervised Multi-Modal Semantic Segmentation

2023-04-21 · Harsh Maheshwari, Yen-Cheng Liu, Zsolt Kira

Using multiple spatial modalities has been proven helpful in improving semantic segmentation performance. However, there are several real-world challenges that have yet to be addressed: (a) improving label efficiency and…

RGBD Semantic SegmentationRobust Semi-Supervised RGBD Semantic SegmentationSegmentationSemantic Segmentation+2

CSFNet: A Cosine Similarity Fusion Network for Real-Time RGB-X Semantic Segmentation of Driving Scenes

2024-07-01 · Danial Qashqai, Emad Mousavian, Shahriar Baradaran Shokouhi, Sattar Mirzakuchaki

Semantic segmentation, as a crucial component of complex visual interpretation, plays a fundamental role in autonomous vehicle vision systems. Recent studies have significantly improved the accuracy of semantic segmentat…

Autonomous VehiclesImage SegmentationReal-Time Semantic SegmentationRGBD Semantic Segmentation+4

Multi-Modal Semantic Inconsistency Detection in Social Media News Posts

2021-05-26 · Scott McCrae, Kehan Wang, Avideh Zakhor

As computer-generated content and deepfakes make steady improvements, semantic approaches to multimedia forensics will become more important. In this paper, we introduce a novel classification architecture for identifyin…

object-detectionObject Detection

SDGOCC: Semantic and Depth-Guided Bird's-Eye View Transformation for 3D Multimodal Occupancy Prediction

2025-07-22 · Zaipeng Duan, Chenxu Dang, Xuzhong Hu, Pei An 외 arxiv

Multimodal 3D occupancy prediction has garnered significant attention for its potential in autonomous driving. However, most existing approaches are single-modality: camera-based methods lack depth information, while LiD…

Autonomous DrivingDepth Estimation