paper-with-me

홈 › Papers

AWM-Fuse: Multi-Modality Image Fusion for Adverse Weather via Global and Local Text Perception

2025-08-23 · Xilai Li, Huichun Liu, Xiaosong Li, Tao Ye, Zhenyu Kuang, Huafeng Li arxiv

Multi-modality image fusion (MMIF) in adverse weather aims to address the loss of visual information caused by weather-related degradations, providing clearer scene representations. Although less studies have attempted to incorporate textual information to improve semantic perception, they often lack effective categorization and thorough analysis of textual content. In response, we propose AWM-Fuse, a novel fusion method for adverse weather conditions, designed to handle multiple degradations through global and local text perception within a unified, shared weight architecture. In particular, a global feature perception module leverages BLIP-produced captions to extract overall scene features and identify primary degradation types, thus promoting generalization across various adverse weather conditions. Complementing this, the local module employs detailed scene descriptions produced by ChatGPT to concentrate on specific degradation effects through concrete textual cues, thereby capturing finer details. Furthermore, textual descriptions are used to constrain the generation of fusion images, effectively steering the network learning process toward better alignment with real semantic labels, thereby promoting the learning of more meaningful visual features. Extensive experiments demonstrate that AWM-Fuse outperforms current state-of-the-art methods in complex weather conditions and downstream tasks. Our code is available at https://github.com/Feecuin/AWM-Fuse.

📄 PDF Abstract BibTeX arXiv:2508.16881

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multi-modality Image Fusion under Adverse Weather: Mask-Guided Feature Restoration and Interaction

2026-06-25 · Xilai Li, Xiaosong Li, Haishu Tan, Tao Ye 외 arxiv

Multi-modality image fusion (MMIF) enhances scene representation by exploiting complementary cues from different modalities. Adverse weather, however, causes significant image degradation, disrupting feature representati…

Representation Learning

Pay "Attention" to Adverse Weather: Weather-aware Attention-based Object Detection

2022-04-22 · Saket S. Chaturvedi, Lan Zhang, Xiaoyong Yuan

Despite the recent advances of deep neural networks, object detection for adverse weather remains challenging due to the poor perception of some sensors in adverse weather. Instead of relying on one single sensor, multim…

object-detectionObject Detection

CDDFuse: Correlation-Driven Dual-Branch Feature Decomposition for Multi-Modality Image Fusion

2022-11-26 · CVPR 2023 1 · Zixiang Zhao, Haowen Bai, Jiangshe Zhang, Yulun Zhang 외

Multi-modality (MM) image fusion aims to render fused images that maintain the merits of different modalities, e.g., functional highlight and detailed textures. To tackle the challenge in modeling cross-modality features…

object-detectionObject DetectionSemantic Segmentation

MambaDFuse: A Mamba-based Dual-phase Model for Multi-modality Image Fusion

2024-04-12 · Zhe Li, Haiwei Pan, Kejia Zhang, Yuhua Wang 외

Multi-modality image fusion (MMIF) aims to integrate complementary information from different modalities into a single fused image to represent the imaging scene and facilitate downstream visual tasks comprehensively. In…

Image ReconstructionMambaobject-detectionObject Detection

Bridging Spectral-wise and Multi-spectral Depth Estimation via Geometry-guided Contrastive Learning

2025-03-02 · Ukcheol Shin, Kyunghyun Lee, Jean Oh

Deploying depth estimation networks in the real world requires high-level robustness against various adverse conditions to ensure safe and reliable autonomy. For this purpose, many autonomous vehicles employ multi-modal …

Autonomous VehiclesContrastive LearningDepth Estimation