paper-with-me

홈 › Papers

PG-ControlNet: A Physics-Guided ControlNet for Generative Spatially Varying Image Deblurring

2025-11-26 · Hakki Motorcu, Mujdat Cetin arxiv

Spatially varying image deblurring remains a fundamentally ill-posed problem, especially when degradations arise from complex mixtures of motion and other forms of blur under significant noise. State-of-the-art learning-based approaches generally fall into two paradigms: model-based deep unrolling methods that enforce physical constraints by modeling the degradations, but often produce over-smoothed, artifact-laden textures, and generative models that achieve superior perceptual quality yet hallucinate details due to weak physical constraints. In this paper, we propose a novel framework that uniquely reconciles these paradigms by taming a powerful generative prior with explicit, dense physical constraints. Rather than oversimplifying the degradation field, we model it as a dense continuum of high-dimensional compressed kernels, ensuring that minute variations in motion and other degradation patterns are captured. We leverage this rich descriptor field to condition a ControlNet architecture, strongly guiding the diffusion sampling process. Extensive experiments demonstrate that our method effectively bridges the gap between physical accuracy and perceptual realism, outperforming state-of-the-art model-based methods as well as generative baselines in challenging, severely blurred scenarios.

📄 PDF Abstract BibTeX arXiv:2511.21043

Code (0)

등록된 구현이 없습니다.

Tasks

Image Deblurring

Similar Papers 제목 키워드 기반

FineControlNet: Fine-level Text Control for Image Generation with Spatially Aligned Text Control Injection

2023-12-14 · Hongsuk Choi, Isaac Kasahara, Selim Engin, Moritz Graule 외

Recently introduced ControlNet has the ability to steer the text-driven image generation process with geometric input such as human 2D pose, or edge features. While ControlNet provides control over the geometric form of …

Image Generation

VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet

2023-07-26 · Zhihao Hu, Dong Xu

Recently, diffusion models like StableDiffusion have achieved impressive image generation results. However, the generation process of such diffusion models is uncontrollable, which makes it hard to generate videos with c…

Image Generation

Cocktail: Mixing Multi-Modality Controls for Text-Conditional Image Generation

2023-06-01 · Minghui Hu, Jianbin Zheng, Daqing Liu, Chuanxia Zheng 외

Text-conditional diffusion models are able to generate high-fidelity images with diverse contents. However, linguistic representations frequently exhibit ambiguous descriptions of the envisioned objective imagery, requir…

Conditional Image GenerationImage Generation

Cocktail: Mixing Multi-Modality Control for Text-Conditional Image Generation

2023-09-21 · NeurIPS 2023 11

Text-conditional diffusion models are able to generate high-fidelity images with diverse contents. However, linguistic representations frequently exhibit ambiguous descriptions of the envisioned objective imagery, requir…

SpecMaskFoley: Steering Pretrained Spectral Masked Generative Transformer Toward Synchronized Video-to-audio Synthesis via ControlNet

2025-05-22 · Zhi Zhong, Akira Takahashi, Shuyang Cui, Keisuke Toyama 외

Foley synthesis aims to synthesize high-quality audio that is both semantically and temporally aligned with video frames. Given its broad application in creative industries, the task has gained increasing attention in th…

Audio Synthesis