paper-with-me

홈 › Papers

MAFNet:Multi-frequency Adaptive Fusion Network for Real-time Stereo Matching

2025-12-04 · Ao Xu, Rujin Zhao, Xiong Xu, Boceng Huang, Yujia Jia, Hongfeng Long, Fuxuan Chen, Zilong Cao, Fangyuan Chen arxiv

Existing stereo matching networks typically rely on either cost-volume construction based on 3D convolutions or deformation methods based on iterative optimization. The former incurs significant computational overhead during cost aggregation, whereas the latter often lacks the ability to model non-local contextual information. These methods exhibit poor compatibility on resource-constrained mobile devices, limiting their deployment in real-time applications. To address this, we propose a Multi-frequency Adaptive Fusion Network (MAFNet), which can produce high-quality disparity maps using only efficient 2D convolutions. Specifically, we design an adaptive frequency-domain filtering attention module that decomposes the full cost volume into high-frequency and low-frequency volumes, performing frequency-aware feature aggregation separately. Subsequently, we introduce a Linformer-based low-rank attention mechanism to adaptively fuse high- and low-frequency information, yielding more robust disparity estimation. Extensive experiments demonstrate that the proposed MAFNet significantly outperforms existing real-time methods on public datasets such as Scene Flow and KITTI 2015, showing a favorable balance between accuracy and real-time performance.

📄 PDF Abstract BibTeX arXiv:2512.04358

Code (0)

등록된 구현이 없습니다.

Tasks

Disparity Estimation

Similar Papers 제목 키워드 기반

Multi-scale Adaptive Fusion Network for Hyperspectral Image Denoising

2023-04-19 · Haodong Pan, Feng Gao, Junyu Dong, Qian Du

Removing the noise and improving the visual quality of hyperspectral images (HSIs) is challenging in academia and industry. Great efforts have been made to leverage local, global or spectral context information for HSI d…

DenoisingHyperspectral Image DenoisingImage Denoising

Cross-Modal Object Tracking via Modality-Aware Fusion Network and A Large-Scale Dataset

2023-12-22 · Lei Liu, Mengya Zhang, Cheng Li, Chenglong Li 외

Visual tracking often faces challenges such as invalid targets and decreased performance in low-light conditions when relying solely on RGB image sequences. While incorporating additional modalities like depth and infrar…

Object TrackingVisual Tracking

MAFNet: A Multi-Attention Fusion Network for RGB-T Crowd Counting

2022-08-14 · PengYu Chen, Junyu Gao, Yuan Yuan, Qi Wang

RGB-Thermal (RGB-T) crowd counting is a challenging task, which uses thermal images as complementary information to RGB images to deal with the decreased performance of unimodal RGB-based methods in scenes with low-illum…

Crowd Counting

Multi-level Attention Fusion Network for Audio-visual Event Recognition

2021-06-12 · Mathilde Brousmiche, Jean Rouat, Stéphane Dupont

Event classification is inherently sequential and multimodal. Therefore, deep neural models need to dynamically focus on the most relevant time window and/or modality of a video. In this study, we propose the Multi-level…

Cross-Modal Purification and Fusion for Small-Object RGB-D Transmission-Line Defect Detection

2026-02-02 · Jiaming Cui, Wenqiang Li, Shuai Zhou, Ruifeng Qin 외 arxiv

Transmission line defect detection remains challenging for automated UAV inspection due to the dominance of small-scale defects, complex backgrounds, and illumination variations. Existing RGB-based detectors, despite rec…