paper-with-me

홈 › Papers

CtrlFuse: Mask-Prompt Guided Controllable Infrared and Visible Image Fusion

2026-01-12 · Yiming Sun, Yuan Ruan, Qinghua Hu, Pengfei Zhu arxiv

Infrared and visible image fusion generates all-weather perception-capable images by combining complementary modalities, enhancing environmental awareness for intelligent unmanned systems. Existing methods either focus on pixel-level fusion while overlooking downstream task adaptability or implicitly learn rigid semantics through cascaded detection/segmentation models, unable to interactively address diverse semantic target perception needs. We propose CtrlFuse, a controllable image fusion framework that enables interactive dynamic fusion guided by mask prompts. The model integrates a multi-modal feature extractor, a reference prompt encoder (RPE), and a prompt-semantic fusion module (PSFM). The RPE dynamically encodes task-specific semantic prompts by fine-tuning pre-trained segmentation models with input mask guidance, while the PSFM explicitly injects these semantics into fusion features. Through synergistic optimization of parallel segmentation and fusion branches, our method achieves mutual enhancement between task performance and fusion quality. Experiments demonstrate state-of-the-art results in both fusion controllability and segmentation accuracy, with the adapted task branch even outperforming the original segmentation model.

📄 PDF Abstract BibTeX arXiv:2601.08619

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ConFusion: Continuous Fusion Space Learning for Fine-Grained Controllable Infrared and Visible Image Fusion

2026-07-26 · Guo Yurong, He Yufei, Li Yonghao, Chang Dongliang 외 arxiv

Controllable infrared-visible image fusion aims to integrate complementary thermal and structural information with flexible region-aware modulation, producing fused images that adapt to diverse user requirements and down…

Beyond Full Labels: Energy-Double-Guided Single-Point Prompt for Infrared Small Target Label Generation

2024-08-15 · Shuai Yuan, Hanlin Qin, Renke Kou, Xiang Yan 외

We pioneer a learning-based single-point prompt paradigm for infrared small target label generation (IRSTLG) to lobber annotation burdens. Unlike previous clustering-based methods, our intuition is that point-guided mask…

Pseudo Label

Prompt-Aware Controllable Shadow Removal

2025-01-25 · Kerui Chen, Zhiliang Wu, Wenjin Hou, Kun Li 외

Shadow removal aims to restore the image content in shadowed regions. While deep learning-based methods have shown promising results, they still face key challenges: 1) uncontrolled removal of all shadows, or 2) controll…

Shadow Removal

Mask What Matters: Controllable Text-Guided Masking for Self-Supervised Medical Image Analysis

2025-09-27 · Ruilang Wang, Shuotong Xu, Bowen Liu, Runlin Huang 외 arxiv

The scarcity of annotated data in specialized domains such as medical imaging presents significant challenges to training robust vision models. While self-supervised masked image modeling (MIM) offers a promising solutio…

Self-Supervised LearningRepresentation Learning

SAIST: Segment Any Infrared Small Target Model Guided by Contrastive Language-Image Pretraining

2025-01-01 · CVPR 2025 1 · Mingjin Zhang, Xiaolong Li, Fei Gao, Jie Guo 외

Infrared Small Target Detection (IRSTD) aims to identify low signal-to-noise ratio small targets in infrared images with complex backgrounds, which is crucial for various applications. However, existing IRSTD methods…

Scene Recognition