paper-with-me

Papers

StructSR: Refuse Spurious Details in Real-World Image Super-Resolution

2025-01-10 · Yachao Li, Dong Liang, Tianyu Ding, Sheng-Jun Huang

Diffusion-based models have shown great promise in real-world image super-resolution (Real-ISR), but often generate content with structural errors and spurious texture details due to the empirical priors and illusions of these models. To address this issue, we introduce StructSR, a simple, effective, and plug-and-play method that enhances structural fidelity and suppresses spurious details for diffusion-based Real-ISR. StructSR operates without the need for additional fine-tuning, external model priors, or high-level semantic knowledge. At its core is the Structure-Aware Screening (SAS) mechanism, which identifies the image with the highest structural similarity to the low-resolution (LR) input in the early inference stage, allowing us to leverage it as a historical structure knowledge to suppress the generation of spurious details. By intervening in the diffusion inference process, StructSR seamlessly integrates with existing diffusion-based Real-ISR models. Our experimental results demonstrate that StructSR significantly improves the fidelity of structure and texture, improving the PSNR and SSIM metrics by an average of 5.27% and 9.36% on a synthetic dataset (DIV2K-Val) and 4.13% and 8.64% on two real-world datasets (RealSR and DRealSR) when integrated with four state-of-the-art diffusion-based Real-ISR methods.

📄 PDF Abstract BibTeX arXiv:2501.05777

Code (1)

lycexe/structsr 공식 구현 pytorch

Tasks

Image Super-ResolutionSSIMSuper-Resolution

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Which Concepts to Forget and How to Refuse? Decomposing Concepts for Continual Unlearning in Large Vision-Language Models

2026-03-23 · Hyundong Jin, Dongyoon Han, Eunwoo Kim arxiv

Continual unlearning poses the challenge of enabling large vision-language models to selectively refuse specific image-instruction pairs in response to sequential deletion requests, while preserving general utility. Howe…

Toxicity Detection for Free

2024-05-29 · Zhanhao Hu, Julien Piet, Geng Zhao, Jiantao Jiao 외

Current LLMs are generally aligned to follow safety requirements and tend to refuse toxic prompts. However, LLMs can fail to refuse toxic prompts or be overcautious and refuse benign examples. In addition, state-of-the-a…

Safety Mirage: How Spurious Correlations Undermine VLM Safety Fine-tuning

2025-03-14 · YiWei Chen, Yuguang Yao, Yihua Zhang, Bingquan Shen 외

Recent vision-language models (VLMs) have made remarkable strides in generative modeling with multimodal inputs, particularly text and images. However, their susceptibility to generating harmful content when exposed to u…

Machine Unlearning

Shortcut to Nowhere: Demystifying Deep Spurious Regression

2026-06-01 · Guanrong Xu, Jessica Li, Hao Wang, Yuzhe Yang arxiv

Real-world regression often exhibits shortcuts: attributes that are spuriously correlated with continuous targets in training, yet unreliable under deployment shifts; regressing targets using such shortcuts may fail cata…

A Detail Based Method for Linear Full Reference Image Quality Prediction

2017-09-10 · Elio D. Di Claudio, Giovanni Jacovitti

In this paper, a novel Full Reference method is proposed for image quality assessment, using the combination of two separate metrics to measure the perceptually distinct impact of detail losses and of spurious details. T…

Image Quality Assessment