paper-with-me

Papers

TriFusion-SR: Joint Tri-Modal Medical Image Fusion and SR

2026-03-10 · Fayaz Ali Dharejo, Sharif S. M. A., Aiman Khalil, Nachiket Chaudhary, Rizwan Ali Naqvi, Radu Timofte arxiv

Multimodal medical image fusion facilitates comprehensive diagnosis by aggregating complementary structural and functional information, but its effectiveness is limited by resolution degradation and modality discrepancies. Existing approaches typically perform image fusion and super-resolution (SR) in separate stages, leading to artifacts and degraded perceptual quality. These limitations are further amplified in tri-modal settings that combine anatomical modalities (e.g., MRI, CT) with functional scans (e.g., PET, SPECT) due to pronounced frequency domain imbalances. We propose TriFusionSR, a wavelet-guided conditional diffusion framework for joint tri-modal fusion and SR. The framework explicitly decomposes multimodal features into frequency bands using the 2D Discrete Wavelet Transform, enabling frequency-aware crossmodal interaction. We further introduce a Rectified Wavelet Features (RWF) strategy for latent coefficient calibration, followed by an Adaptive Spatial-Frequency Fusion (ASFF) module with gated channel-spatial attention to enable structure-driven multimodal refinement. Extensive experiments demonstrate state-of-the-art performance, achieving 4.8-12.4% PSNR improvement and substantial reductions in RMSE and LPIPS across multiple upsampling scales.

📄 PDF Abstract BibTeX arXiv:2603.09702

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TriFusion-AE: Language-Guided Depth and LiDAR Fusion for Robust Point Cloud Processing

2025-09-23 · Susmit Neogi arxiv

LiDAR-based perception is central to autonomous driving and robotics, yet raw point clouds remain highly vulnerable to noise, occlusion, and adversarial corruptions. Autoencoders offer a natural framework for denoising a…

Representation LearningAutonomous DrivingPoint Clouds

DistriFusion: Distributed Parallel Inference for High-Resolution Diffusion Models

2024-02-29 · CVPR 2024 1 · Muyang Li, Tianle Cai, Jiaxin Cao, Qinsheng Zhang 외

Diffusion models have achieved great success in synthesizing high-quality images. However, generating high-resolution images with diffusion models is still challenging due to the enormous computational costs, resulting i…

GPU

Partially Conditioned Patch Parallelism for Accelerated Diffusion Model Inference

2024-12-04 · XiuYu Zhang, Zening Luo, Michelle E. Lu

Diffusion models have exhibited exciting capabilities in generating images and are also very promising for video creation. However, the inference speed of diffusion models is limited by the slow sampling process, restric…

DenoisingImage Generation

Discrete Diffusion Models with MLLMs for Unified Medical Multimodal Generation

2025-10-07 · Jiawei Mao, Yuhan Wang, Lifeng Chen, Can Zhao 외 arxiv

Recent advances in generative medical models are constrained by modality-specific scenarios that hinder the integration of complementary evidence from imaging, pathology, and clinical notes. This fragmentation limits the…

multimodal generation

MedPatch: Confidence-Guided Multi-Stage Fusion for Multimodal Clinical Data

2025-08-07 · Baraa Al Jorf, Farah Shamout arxiv

Clinical decision-making relies on the integration of information across various data modalities, such as clinical time-series, medical images and textual reports. Compared to other domains, real-world medical data is he…

Mortality Prediction