paper-with-me

Papers

Multispectral Video Semantic Segmentation: A Benchmark Dataset and Baseline

2023-01-01 · CVPR 2023 1 · Wei Ji, Jingjing Li, Cheng Bian, Zongwei Zhou, Jiaying Zhao, Alan L. Yuille, Li Cheng

Robust and reliable semantic segmentation in complex scenes is crucial for many real-life applications such as autonomous safe driving and nighttime rescue. In most approaches, it is typical to make use of RGB images as input. They however work well only in preferred weather conditions; when facing adverse conditions such as rainy, overexposure, or low-light, they often fail to deliver satisfactory results. This has led to the recent investigation into multispectral semantic segmentation, where RGB and thermal infrared (RGBT) images are both utilized as input. This gives rise to significantly more robust segmentation of image objects in complex scenes and under adverse conditions. Nevertheless, the present focus in single RGBT image input restricts existing methods from well addressing dynamic real-world scenes. Motivated by the above observations, in this paper, we set out to address a relatively new task of semantic segmentation of multispectral video input, which we refer to as Multispectral Video Semantic Segmentation, or MVSS in short. An in-house MVSeg dataset is thus curated, consisting of 738 calibrated RGB and thermal videos, accompanied by 3,545 fine-grained pixel-level semantic annotations of 26 categories. Our dataset contains a wide range of challenging urban scenes in both daytime and nighttime. Moreover, we propose an effective MVSS baseline, dubbed MVNet, which is to our knowledge the first model to jointly learn semantic representations from multispectral and temporal contexts. Comprehensive experiments are conducted using various semantic segmentation models on the MVSeg dataset. Empirically, the engagement of multispectral video input is shown to lead to significant improvement in semantic segmentation; the effectiveness of our MVNet baseline has also been verified.

📄 PDF Abstract BibTeX

Code (1)

jiwei0921/MVSS-Baseline 공식 구현 pytorch

Tasks

SegmentationSemantic SegmentationVideo Semantic Segmentation

Methods 이 논문이 사용한 방법론

fail 설명 없음

Similar Papers 제목 키워드 기반

High-Resolution Multispectral Dataset for Semantic Segmentation

2017-03-06 · Ronald Kemker, Carl Salvaggio, Christopher Kanan

Unmanned aircraft have decreased the cost required to collect remote sensing imagery, which has enabled researchers to collect high-spatial resolution data from multiple sensor modalities more frequently and easily. The …

General ClassificationSemantic SegmentationVocal Bursts Intensity Prediction

Multispectral Pedestrian Detection via Simultaneous Detection and Segmentation

2018-08-14 · Chengyang Li, Dan Song, Ruofeng Tong, Min Tang

Multispectral pedestrian detection has attracted increasing attention from the research community due to its crucial competence for many around-the-clock applications (e.g., video surveillance and autonomous driving), es…

Autonomous DrivingMultispectral Object DetectionPedestrian DetectionSemantic Segmentation

Online Mutual Foreground Segmentation for Multispectral Stereo Videos

2018-09-08 · Pierre-Luc St-Charles, Guillaume-Alexandre Bilodeau, Robert Bergevin

The segmentation of video sequences into foreground and background regions is a low-level process commonly used in video content analysis and smart surveillance applications. Using a multispectral camera setup can improv…

Foreground Segmentation

Enhancing Environmental Monitoring through Multispectral Imaging: The WasteMS Dataset for Semantic Segmentation of Lakeside Waste

2024-07-24 · Qinfeng Zhu, Ningxin Weng, Lei Fan, Yuanzhi Cai

Environmental monitoring of lakeside green areas is crucial for environmental protection. Compared to manual inspections, computer vision technologies offer a more efficient solution when deployed on-site. Multispectral …

SegmentationSemantic Segmentation

Fusion of Multispectral Data Through Illumination-aware Deep Neural Networks for Pedestrian Detection

2018-02-27 · Dayan Guan, Yanpeng Cao, Jun Liang, Yanlong Cao 외

Multispectral pedestrian detection has received extensive attention in recent years as a promising solution to facilitate robust human target detection for around-the-clock applications (e.g. security surveillance and au…

Autonomous DrivingMultispectral Object DetectionMulti-Task LearningPedestrian Detection+1