paper-with-me

홈 › Papers

FMRFusion: Frequency-Aware Multi-View Representation Learning for Heterogeneous Image Fusion

2026-06-06 · Tao Zhoua, Yunlong Liu, Qinghui Chen, Zekai Zhang, Minlong Sun, Changlin Biana, Dagang Li, Wenmin Wang, Jinglin Zhang arxiv

Infrared and visible image fusion aims to generate a composite image that retains significant target information and preserves detailed textures, integrating two heterogeneous modalities. Previous image fusion methods typically adopt a single-module stacking approach to extract features from the two modalities. However, these approaches may result in incomplete learning of their distinct characteristics, thereby limiting the fusion effectiveness and constrain ing robustness in real-world heterogeneous data scenarios. To address these challenges, we propose FMRFusion, a frequency-aware multi-view representation learning network for Heterogeneous Image Fusion. A Multi-Scale Struc tural Perception Module is introduced to effectively capture discriminative structures, extracting fine-grained local structures and essential contextual information. A bilinear frequency decomposition mechanism is employed to sepa rate features into high-frequency and low-frequency components, enabling joint modeling of local details and global representations across different frequency domains. Moreover, a Cross-View Complementary Interaction is incorpo rated to explicitly model and fuse the complementary characteristics between reflected light information and radiative intensity responses, facilitating effective cross-view interaction. We further improve the Performance of the fused results by flow matching, which progressively refines the fused features by learning the transformation from coarse data to high-quality representations. Extensive experiments conducted on multiple benchmark datasets demonstrate that FMRFusion achieves superior and consistent performance across a range of fusion tasks, especially in nighttime scenarios

📄 PDF Abstract BibTeX arXiv:2606.07985

Code (0)

등록된 구현이 없습니다.

Tasks

Representation Learning

Similar Papers 제목 키워드 기반

ConFi-GS Confidence-Guided High-Frequency Injection for 3D Gaussian Splatting Super-Resolution

2026-05-24 · Jiaxiang Li, Zongtan Zhou, Zhen Tan, Yadong Liu 외 arxiv

Reconstructing high-quality 3D scenes from low-resolution multi-view images remains challenging for 3D Gaussian Splatting (3DGS), because insufficient high-frequency observations often lead to blurred textures, weak boun…

MFAF: An EVA02-Based Multi-scale Frequency Attention Fusion Method for Cross-View Geo-Localization

2025-09-16 · YiTong Liu, TianZhu Liu, YanFeng GU arxiv

Cross-view geo-localization aims to determine the geographical location of a query image by matching it against a gallery of images. This task is challenging due to the significant appearance variations of objects observ…

Drone navigation

FLAIR: Frequency- and Locality-Aware Implicit Neural Representations

2025-08-19 · Sukhun Ko, Seokhyun Youn, Dahyeon Kye, Kyle Min 외 arxiv

Implicit Neural Representations (INRs) leverage neural networks to map coordinates to corresponding signals, enabling continuous and compact representations. This paradigm has driven significant advances in various visio…

3D Shape ReconstructionNovel View Synthesis

Faster 3D Gaussian Splatting Convergence via Structure-Aware Densification

2026-04-30 · Linjie Lyu, Ayush Tewari, Jianchun Chen, Thomas Leimkühler 외 arxiv

3D Gaussian Splatting has emerged as a powerful scene representation for real-time novel-view synthesis. However, its standard adaptive density control relies on screen-space positional gradients, which do not distinguis…

LookCloser: Frequency-aware Radiance Field for Tiny-Detail Scene

2025-03-24 · CVPR 2025 1 · XiaoYu Zhang, Weihong Pan, Chong Bao, Xiyu Zhang 외

Humans perceive and comprehend their surroundings through information spanning multiple frequencies. In immersive scenes, people naturally scan their environment to grasp its overall structure while examining fine detail…

NeRF