paper-with-me

홈 › Papers

Bridging Human Evaluation to Infrared and Visible Image Fusion

2026-03-04 · Jinyuan Liu, Xingyuan Li, Qingyun Mei, Haoyuan Xu, Zhiying Jiang, Long Ma, Risheng Liu, Xin Fan arxiv

Infrared and visible image fusion (IVIF) integrates complementary modalities to enhance scene perception. Current methods predominantly focus on optimizing handcrafted losses and objective metrics, often resulting in fusion outcomes that do not align with human visual preferences. This challenge is further exacerbated by the ill-posed nature of IVIF, which severely limits its effectiveness in human perceptual environments such as security surveillance and driver assistance systems. To address these limitations, we propose a feedback reinforcement framework that bridges human evaluation to infrared and visible image fusion. To address the lack of human-centric evaluation metrics and data, we introduce the first large-scale human feedback dataset for IVIF, containing multidimensional subjective scores and artifact annotations, and enriched by a fine-tuned large language model with expert review. Based on this dataset, we design a domain-specific reward function and train a reward model to quantify perceptual quality. Guided by this reward, we fine-tune the fusion network through Group Relative Policy Optimization, achieving state-of-the-art performance that better aligns fused images with human aesthetics. Code is available at https://github.com/ALKA-Wind/EVAFusion.

📄 PDF Abstract BibTeX arXiv:2603.03871

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ROMA: Cross-Domain Region Similarity Matching for Unpaired Nighttime Infrared to Daytime Visible Video Translation

2022-04-26 · Zhenjie Yu, Kai Chen, Shuang Li, Bingfeng Han 외

Infrared cameras are often utilized to enhance the night vision since the visible light cameras exhibit inferior efficacy without sufficient illumination. However, infrared data possesses inadequate color contrast and re…

Translation

Supervised Image Translation from Visible to Infrared Domain for Object Detection

2024-08-03 · Prahlad Anand, Qiranul Saadiyean, Aniruddh Sikdar, Nalini N 외

This study aims to learn a translation from visible to infrared imagery, bridging the domain gap between the two modalities so as to improve accuracy on downstream tasks including object detection. Previous approaches at…

Generative Adversarial NetworkObjectobject-detectionObject Detection+2

IV-tuning: Parameter-Efficient Transfer Learning for Infrared-Visible Tasks

2024-12-21 · Yaming Zhang, Chenqiang Gao, Fangcen Liu, Junjie Guo 외

Infrared-visible (IR-VIS) tasks, such as semantic segmentation and object detection, greatly benefit from the advantage of combining infrared and visible modalities. To inherit the general representations of the Vision F…

object-detectionObject DetectionSemantic SegmentationTransfer Learning

A Joint Convolution Auto-encoder Network for Infrared and Visible Image Fusion

2022-01-26 · Zhancheng Zhang, Yuanhao Gao, Mengyu Xiong, Xiaoqing Luo 외

Background: Leaning redundant and complementary relationships is a critical step in the human visual system. Inspired by the infrared cognition ability of crotalinae animals, we design a joint convolution auto-encoder (J…

DecoderInfrared And Visible Image Fusion

Bridging the Gap: Multi-Level Cross-Modality Joint Alignment for Visible-Infrared Person Re-Identification

2023-07-17 · Tengfei Liang, Yi Jin, Wu Liu, Tao Wang 외

Visible-Infrared person Re-IDentification (VI-ReID) is a challenging cross-modality image retrieval task that aims to match pedestrians' images across visible and infrared cameras. To solve the modality gap, existing mai…

Cross-Modality Person Re-identificationimage-classificationImage ClassificationImage Retrieval+3