paper-with-me

Papers

Multimodal Material Segmentation

2022-01-01 · CVPR 2022 1 · Yupeng Liang, Ryosuke Wakaki, Shohei Nobuhara, Ko Nishino

Recognition of materials from their visual appearance is essential for computer vision tasks, especially those that involve interaction with the real world. Material segmentation, i.e., dense per-pixel recognition of materials, remains challenging as, unlike objects, materials do not exhibit clearly discernible visual signatures in their regular RGB appearances. Different materials, however, do lead to different radiometric behaviors, which can often be captured with non-RGB imaging modalities. We realize multimodal material segmentation from RGB, polarization, and near-infrared images. We introduce the MCubeS dataset (from MultiModal Material Segmentation) which contains 500 sets of multimodal images capturing 42 street scenes. Ground truth material segmentation as well as semantic segmentation are annotated for every image and pixel. We also derive a novel deep neural network, MCubeSNet, which learns to focus on the most informative combinations of imaging modalities for each material class with a newly derived region-guided filter selection (RGFS) layer. We use semantic segmentation, as a prior to "guide" this filter selection. To the best of our knowledge, our work is the first comprehensive study on truly multimodal material segmentation. We believe our work opens new avenues of practical use of material information in safety critical applications.

📄 PDF Abstract BibTeX

Code (1)

kyotovision-public/multimodal-material-segmentation 공식 구현 pytorch

Tasks

Material SegmentationSegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

MMSFormer: Multimodal Transformer for Material and Semantic Segmentation

2023-09-07 · Md Kaykobad Reza, Ashley Prater-Bennette, M. Salman Asif

Leveraging information across diverse modalities is known to enhance performance on multimodal segmentation tasks. However, effectively fusing information from different modalities remains challenging due to the unique c…

SegmentationSemantic SegmentationThermal Image Segmentation

Bidirectional Cross-Attention Fusion of High-Resolution RGB and Low-Resolution Hyperspectral Inputs for Multimodal Semantic Segmentation

2026-03-14 · Jonas V. Funk, Lukas Roming, Andreas Michel, Paul Bäcker 외 arxiv

Multimodal semantic segmentation with heterogeneous sensors must reconcile complementary information across modalities that differ in spatial resolution and channel dimensionality. In particular, high-resolution RGB imag…

Semantic Segmentation

LiveSeg: Unsupervised Multimodal Temporal Segmentation of Long Livestream Videos

2022-10-12 · JieLin Qiu, Franck Dernoncourt, Trung Bui, Zhaowen Wang 외

Livestream videos have become a significant part of online learning, where design, digital marketing, creative painting, and other skills are taught by experienced experts in the sessions, making them valuable materials.…

MarketingSegmentation

EDEN: Multimodal Synthetic Dataset of Enclosed GarDEN Scenes

2020-11-09 · Hoang-An Le, Thomas Mensink, Partha Das, Sezer Karaoglu 외

Multimodal large-scale datasets for outdoor scenes are mostly designed for urban driving problems. The scenes are highly structured and semantically different from scenarios seen in nature-centered scenes such as gardens…

Depth EstimationDepth PredictionOptical Flow EstimationSegmentation+1

Material Segmentation of Multi-View Satellite Imagery

2019-04-17 · Matthew Purri, Jia Xue, Kristin Dana, Matthew Leotta 외

Material recognition methods use image context and local cues for pixel-wise classification. In many cases only a single image is available to make a material prediction. Image sequences, routinely acquired in applicatio…

Material RecognitionMaterial SegmentationSegmentationSemantic Segmentation