paper-with-me

홈 › Papers

Boosting ViT-based MRI Reconstruction from the Perspectives of Frequency Modulation, Spatial Purification, and Scale Diversification

2024-12-14 · Yucong Meng, Zhiwei Yang, Yonghong Shi, Zhijian Song

The accelerated MRI reconstruction process presents a challenging ill-posed inverse problem due to the extensive under-sampling in k-space. Recently, Vision Transformers (ViTs) have become the mainstream for this task, demonstrating substantial performance improvements. However, there are still three significant issues remain unaddressed: (1) ViTs struggle to capture high-frequency components of images, limiting their ability to detect local textures and edge information, thereby impeding MRI restoration; (2) Previous methods calculate multi-head self-attention (MSA) among both related and unrelated tokens in content, introducing noise and significantly increasing computational burden; (3) The naive feed-forward network in ViTs cannot model the multi-scale information that is important for image restoration. In this paper, we propose FPS-Former, a powerful ViT-based framework, to address these issues from the perspectives of frequency modulation, spatial purification, and scale diversification. Specifically, for issue (1), we introduce a frequency modulation attention module to enhance the self-attention map by adaptively re-calibrating the frequency information in a Laplacian pyramid. For issue (2), we customize a spatial purification attention module to capture interactions among closely related tokens, thereby reducing redundant or irrelevant feature representations. For issue (3), we propose an efficient feed-forward network based on a hybrid-scale fusion strategy. Comprehensive experiments conducted on three public datasets show that our FPS-Former outperforms state-of-the-art methods while requiring lower computational costs.

📄 PDF Abstract BibTeX arXiv:2412.10776

Code (0)

등록된 구현이 없습니다.

Tasks

Image RestorationMRI Reconstruction

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Dual-domain Modulation Network for Lightweight Image Super-Resolution

2025-03-13 · Wenjie Li, Heng Guo, Yuefeng Hou, Guangwei Gao 외

Lightweight image super-resolution (SR) aims to reconstruct high-resolution images from low-resolution images with limited computational costs. We find existing frequency-based SR methods cannot balance the reconstructio…

Image Super-ResolutionSuper-Resolution

HFGS: 4D Gaussian Splatting with Emphasis on Spatial and Temporal High-Frequency Components for Endoscopic Scene Reconstruction

2024-05-28 · Haoyu Zhao, Xingyue Zhao, Lingting Zhu, Weixi Zheng 외

Robot-assisted minimally invasive surgery benefits from enhancing dynamic scene reconstruction, as it improves surgical outcomes. While Neural Radiance Fields (NeRF) have been effective in scene reconstruction, their slo…

NeRFNeural Rendering

Rethinking Waveform for 6G: Harnessing Delay-Doppler Alignment Modulation

2024-06-13 · Zhiqiang Xiao, Xianda Liu, Yong Zeng, J. Andrew Zhang 외

Waveform design has served as a cornerstone for each generation of mobile communication systems. The future sixth-generation (6G) mobile communication networks are expected to employ larger-scale antenna arrays and explo…

Generalized code index modulation-aided frequency offset realign multiple-antenna spatial modulation approach for next-generation green communication systems

2024-08-16 · Bang Huang, Jiajie Xu, Mohamed-Slim Alouini

For next-generation green communication systems, this article proposes an innovative communication system based on frequency-diverse array-multiple-input multiple-output (FDA-MIMO) technology, which aims to achieve high …

Adaptive Local Frequency Filtering for Fourier-Encoded Implicit Neural Representations

2026-04-03 · Ligen Shi, Jun Qiu, Yuhang Zheng, Zengyu Pang 외 arxiv

Fourier-encoded implicit neural representations (INRs) have shown strong capability in modeling continuous signals from discrete samples. However, conventional Fourier feature mappings use a fixed set of frequencies over…