paper-with-me

홈 › Papers

Context-Aware Pretraining for Efficient Blind Image Decomposition

2023-01-01 · CVPR 2023 1 · Chao Wang, Zhedong Zheng, Ruijie Quan, Yifan Sun, Yi Yang

In this paper, we study Blind Image Decomposition (BID), which is to uniformly remove multiple types of degradation at once without foreknowing the noise type. There remain two practical challenges: (1) Existing methods typically require massive data supervision, making them infeasible to real-world scenarios. (2) The conventional paradigm usually focuses on mining the abnormal pattern of a superimposed image to separate the noise, which de facto conflicts with the primary image restoration task. Therefore, such a pipeline compromises repairing efficiency and authenticity. In an attempt to solve the two challenges in one go, we propose an efficient and simplified paradigm, called Context-aware Pretraining (CP), with two pretext tasks: mixed image separation and masked image reconstruction. Such a paradigm reduces the annotation demands and explicitly facilitates context-aware feature learning. Assuming the restoration process follows a structure-to-texture manner, we also introduce a Context-aware Pretrained network (CPNet). In particular, CPNet contains two transformer-based parallel encoders, one information fusion module, and one multi-head prediction module. The information fusion module explicitly utilizes the mutual correlation in the spatial-channel dimension, while the multi-head prediction module facilitates texture-guided appearance flow. Moreover, a new sampling loss along with an attribute label constraint is also deployed to make use of the spatial context, leading to high-fidelity image restoration. Extensive experiments on both real and synthetic benchmarks show that our method achieves competitive performance for various BID tasks.

📄 PDF Abstract BibTeX

Code (1)

oliiveralien/cpnet 공식 구현 paddle

Tasks

AttributeImage ReconstructionImage Restoration

Similar Papers 제목 키워드 기반

Strong and Controllable Blind Image Decomposition

2024-03-15 · Zeyu Zhang, Junlin Han, Chenhui Gou, Hongdong Li 외

Blind image decomposition aims to decompose all components present in an image, typically used to restore a multi-degraded input image. While fully recovering the clean image is appealing, in some scenarios, users might …

The Attribution Blind Spot: Detecting When Language Models Rely on Memory Rather Than Retrieved Context

2026-05-26 · Zhe Yu, Wenpeng Xing, Yunzhao Wei, Bo Yang 외 arxiv

Retrieval-augmented generation promises to ground language model outputs in external evidence, yet the field has no reliable way to verify whether retrieved context actually governs generation -- a prerequisite for any h…

The Spatial Blindspot of Vision-Language Models

2026-01-15 · Nahid Alam, Leema Krishna Murali, Siddhant Bharadwaj, Patrick Liu 외 arxiv

Vision-language models (VLMs) have advanced rapidly, but their ability to capture spatial relationships remains a blindspot. Current VLMs are typically built with contrastive language-image pretraining (CLIP) style image…

Spatial Reasoning

Blind Image Decomposition

2021-08-25 · Junlin Han, Weihao Li, Pengfei Fang, Chunyi Sun 외

We propose and study a novel task named Blind Image Decomposition (BID), which requires separating a superimposed image into constituent underlying images in a blind setting, that is, both the source components involved …

Rain Removal

Bridging Component Learning with Degradation Modelling for Blind Image Super-Resolution

2022-12-03 · Yixuan Wu, Feng Li, Huihui Bai, Weisi Lin 외

Convolutional Neural Network (CNN)-based image super-resolution (SR) has exhibited impressive success on known degraded low-resolution (LR) images. However, this type of approach is hard to hold its performance in practi…

Image Super-ResolutionSuper-Resolution