paper-with-me

Papers

Polarization-driven Semantic Segmentation via Efficient Attention-bridged Fusion

2020-11-26 · Kaite Xiang, Kailun Yang, Kaiwei Wang

Semantic Segmentation (SS) is promising for outdoor scene perception in safety-critical applications like autonomous vehicles, assisted navigation and so on. However, traditional SS is primarily based on RGB images, which limits the reliability of SS in complex outdoor scenes, where RGB images lack necessary information dimensions to fully perceive unconstrained environments. As preliminary investigation, we examine SS in an unexpected obstacle detection scenario, which demonstrates the necessity of multimodal fusion. Thereby, in this work, we present EAFNet, an Efficient Attention-bridged Fusion Network to exploit complementary information coming from different optical sensors. Specifically, we incorporate polarization sensing to obtain supplementary information, considering its optical characteristics for robust representation of diverse materials. By using a single-shot polarization sensor, we build the first RGB-P dataset which consists of 394 annotated pixel-aligned RGB-Polarization images. A comprehensive variety of experiments shows the effectiveness of EAFNet to fuse polarization and RGB information, as well as the flexibility to be adapted to other sensor combination scenarios.

📄 PDF Abstract BibTeX arXiv:2011.13313

Code (1)

Katexiang/EAFNet 공식 구현 tf

Tasks

Autonomous VehiclesSemantic Segmentation

Similar Papers 제목 키워드 기반

ShareCMP: Polarization-Aware RGB-P Semantic Segmentation

2023-12-06 · Zhuoyan Liu, Bo wang, Lizhi Wang, Chenyu Mao 외

Multimodal semantic segmentation is developing rapidly, but the modality of RGB-Polarization remains underexplored. To delve into this problem, we construct a UPLight RGB-P segmentation benchmark with 12 typical underwat…

Semantic Segmentation

TAViS: Text-bridged Audio-Visual Segmentation with Foundation Models

2025-06-13 · Ziyang Luo, Nian Liu, Xuguang Yang, Salman Khan 외

Audio-Visual Segmentation (AVS) faces a fundamental challenge of effectively aligning audio and visual modalities. While recent approaches leverage foundation models to address data scarcity, they often rely on single-mo…

cross-modal alignmentSegmentation

BridgeDiff: Bridging Human Observations and Flat-Garment Synthesis for Virtual Try-Off

2026-03-10 · Shuang Liu, Ao Yu, Linkang Cheng, Xiwen Huang 외 arxiv

Virtual try-off (VTOFF) aims to recover canonical flat-garment representations from images of dressed persons for standardized display and downstream virtual try-on. Prior methods often treat VTOFF as direct image transl…

Virtual Try-OffVirtual Try-on

Glass Segmentation Using Intensity and Spectral Polarization Cues

2022-01-01 · CVPR 2022 1 · Haiyang Mei, Bo Dong, Wen Dong, Jiaxi Yang 외

Transparent and semi-transparent materials pose significant challenges for existing scene understanding and segmentation algorithms due to their lack of RGB texture which impedes the extraction of meaningful features…

Camouflaged Object SegmentationScene UnderstandingSegmentationSemantic Segmentation

Segmentation-Driven Monocular Shape from Polarization based on Physical Model

2026-01-08 · Jinyu Zhang, Xu Ma, Weili Chen arxiv

Monocular shape-from-polarization (SfP) leverages the intrinsic relationship between light polarization properties and surface geometry to recover surface normals from single-view polarized images, providing a compact an…