paper-with-me

Papers

Vision-Language Controlled Deep Unfolding for Joint Medical Image Restoration and Segmentation

2026-01-30 · Ping Chen, Zicheng Huang, Xiangming Wang, Yungeng Liu, Bingyu Liang, Haijin Zeng, Yongyong Chen arxiv

We propose VL-DUN, a principled framework for joint All-in-One Medical Image Restoration and Segmentation (AiOMIRS) that bridges the gap between low-level signal recovery and high-level semantic understanding. While standard pipelines treat these tasks in isolation, our core insight is that they are fundamentally synergistic: restoration provides clean anatomical structures to improve segmentation, while semantic priors regularize the restoration process. VL-DUN resolves the sub-optimality of sequential processing through two primary innovations. (1) We formulate AiOMIRS as a unified optimization problem, deriving an interpretable joint unfolding mechanism where restoration and segmentation are mathematically coupled for mutual refinement. (2) We introduce a frequency-aware Mamba mechanism to capture long-range dependencies for global segmentation while preserving the high-frequency textures necessary for restoration. This allows for efficient global context modeling with linear complexity, effectively mitigating the spectral bias of standard architectures. As a pioneering work in the AiOMIRS task, VL-DUN establishes a new state-of-the-art across multi-modal benchmarks, improving PSNR by 0.92 dB and the Dice coefficient by 9.76\%. Our results demonstrate that joint collaborative learning offers a superior, more robust solution for complex clinical workflows compared to isolated task processing. The codes are provided in https://github.com/cipi666/VLDUN.

📄 PDF Abstract BibTeX arXiv:2601.23103

Code (0)

등록된 구현이 없습니다.

Tasks

Image Restoration

Similar Papers 제목 키워드 기반

Combined Dictionary Unfolding Network with Gradient-Adaptive Fidelity for Transferable Multi-Source Fusion

2026-05-01 · Ge Luo, Jun-Jie Huang, Qi Yu, Tianrui Liu 외 arxiv

Deep Unfolding Network-based methods have emerged as effective solutions for multi-source image fusion by combining model-driven iterative optimization with data-driven deep learning. However, most existing deep unfoldin…

Semantic Segmentation

Vision-Language Gradient Descent-driven All-in-One Deep Unfolding Networks

2025-03-21 · CVPR 2025 1 · Haijin Zeng, Xiangming Wang, Yongyong Chen, Jingyong Su 외

Dynamic image degradations, including noise, blur and lighting inconsistencies, pose significant challenges in image restoration, often due to sensor limitations or adverse environmental conditions. Existing Deep Unfoldi…

AllImage RestorationRain Removal

What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction

2024-07-11 · Sunny Panchal, Apratim Bhattacharyya, Guillaume Berger, Antoine Mercier 외

Vision-language models have shown impressive progress in recent years. However, existing models are largely limited to turn-based interactions, where each turn must be stepped (i.e., prompted) by the user. Open-ended, as…

Evaluating the Pre-Dressing Step: Unfolding Medical Garments Via Imitation Learning

2025-07-24 · David Blanco-Mulero, Júlia Borràs, Carme Torras arxiv

Robotic-assisted dressing has the potential to significantly aid both patients as well as healthcare personnel, reducing the workload and improving the efficiency in clinical settings. While substantial progress has been…

Med-Banana: Learning Quality-Controlled Medical Image Editing from Success-and-Failure Trajectories

2025-11-02 · Zhihui Chen, Qingyuan Lei, Kai He, Yanrui Du 외 arxiv

Text-guided medical image editing must satisfy the requested pathology while preserving anatomy, modality-specific appearance, and clinical plausibility. However, existing datasets largely supervise editors with final ac…

Image Editing