paper-with-me

홈 › Papers

Adapting Text-to-Image Generation with Feature Difference Instruction for Generic Image Restoration

2025-01-01 · CVPR 2025 1 · Chao Wang, Hehe Fan, Huichen Yang, Sarvnaz Karimi, Lina Yao, Yi Yang

Diffusion-based Text-to-Image (T2I) models have demonstrated significant potential in image restoration. However, existing models continue to grapple with challenges such as complex training and prompt design. We introduce a new perspective for improving image restoration by injecting knowledge from pretrained vision-language models into current T2I models. We empirically show that the degradation and content representations in BLIP-2 can be linearly separated, providing promising degradation guidance for image restoration. Specifically, the Feature Difference Instruction (FDI) is first extracted by Q-Formers through a simple subtraction operation based on reference image pairs. Then, we propose a multi-scale FDI adapter to decouple the degradation style and corrupted artifacts, and inject the styleflow exclusively into specific blocks through adapter-tuning, thereby preventing noise interference and eschewing the need for cumbersome weight retraining. In this way, we can train various task-specific adapters according to different degradations, achieving rich detail enhancement in the restoration results. Furthermore, the proposed FDI adapters have attractive properties of practical value, such as composability and generalization ability for all-in-one and mixed-degradation restoration. Extensive experiments under various settings demonstrate that our method has promising repairing quality over 10 image restoration tasks and a wide range of other applications.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationImage RestorationText to Image GenerationText-to-Image Generation

Methods 이 논문이 사용한 방법론

Adapter 설명 없음

Similar Papers 제목 키워드 기반

Style Generation: Image Synthesis based on Coarsely Matched Texts

2023-09-08 · Mengyao Cui, Zhe Zhu, Shao-Ping Lu, Yulu Yang

Previous text-to-image synthesis algorithms typically use explicit textual instructions to generate/manipulate images accurately, but they have difficulty adapting to guidance in the form of coarsely matched texts. In th…

Generative Adversarial NetworkImage GenerationSentenceStory Visualization

3-D PET Image Generation with tumour masks using TGAN

2021-11-02 · Robert V Bergen, Jean-Francois Rajotte, Fereshteh Yousefirizi, Ivan S Klyuzhin 외

Training computer-vision related algorithms on medical images for disease diagnosis or image segmentation is difficult due to the lack of training data, labeled samples, and privacy concerns. For this reason, a robust ge…

Image GenerationImage SegmentationSegmentationSemantic Segmentation+1

Feature Difference Makes Sense: A medical image captioning model exploiting feature difference and tag information

2020-07-01 · ACL 2020 6 · Hyeryun Park, Kyungmo Kim, Jooyoung Yoon, Seongkeun Park 외

Medical image captioning can reduce the workload of physicians and save time and expense by automatically generating reports. However, current datasets are small and limited, creating additional challenges for researcher…

Image CaptioningTAG

DFM: Difference Feature Modeling with Text-Guided Gated Contrastive Loss for Remote Sensing Image Change Captioning

2026-06-25 · Yelin Wang, Zijia Song, Chuanguang Yang, Miaoyu Wang 외 arxiv

The primary goal of Remote Sensing Image Change Captioning (RSICC) is to automatically generate descriptions of changes between remote sensing images captured at different time points. Existing models still rely on a sin…

Change Detection

Frame-Difference Guided Dynamic Region Perception for CLIP Adaptation in Text-Video Retrieval

2025-10-21 · Jiaao Yu, Mingjie Han, Tao Gong, Jian Zhang 외 arxiv

With the rapid growth of video data, text-video retrieval technology has become increasingly important in numerous application scenarios such as recommendation and search. Early text-video retrieval methods suffer from t…

Video AlignmentVideo Retrieval