paper-with-me

홈 › Papers

FICGen: Frequency-Inspired Contextual Disentanglement for Layout-driven Degraded Image Generation

2025-09-01 · Wenzhuang Wang, Yifan Zhao, Mingcan Ma, Ming Liu, Zhonglin Jiang, Yong Chen, Jia Li arxiv

Layout-to-image (L2I) generation has exhibited promising results in natural domains, but suffers from limited generative fidelity and weak alignment with user-provided layouts when applied to degraded scenes (i.e., low-light, underwater). We primarily attribute these limitations to the "contextual illusion dilemma" in degraded conditions, where foreground instances are overwhelmed by context-dominant frequency distributions. Motivated by this, our paper proposes a new Frequency-Inspired Contextual Disentanglement Generative (FICGen) paradigm, which seeks to transfer frequency knowledge of degraded images into the latent diffusion space, thereby facilitating the rendering of degraded instances and their surroundings via contextual frequency-aware guidance. To be specific, FICGen consists of two major steps. Firstly, we introduce a learnable dual-query mechanism, each paired with a dedicated frequency resampler, to extract contextual frequency prototypes from pre-collected degraded exemplars in the training set. Secondly, a visual-frequency enhanced attention is employed to inject frequency prototypes into the degraded generation process. To alleviate the contextual illusion and attribute leakage, an instance coherence map is developed to regulate latent-space disentanglement between individual instances and their surroundings, coupled with an adaptive spatial-frequency aggregation module to reconstruct spatial-frequency mixed degraded representations. Extensive experiments on 5 benchmarks involving a variety of degraded scenarios-from severe low-light to mild blur-demonstrate that FICGen consistently surpasses existing L2I methods in terms of generative fidelity, alignment and downstream auxiliary trainability.

📄 PDF Abstract BibTeX arXiv:2509.01107

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

The Layout Generation Algorithm of Graphic Design Based on Transformer-CVAE

2021-10-08 · Mengxi Guo, Dangqing Huang, Xiaodong Xie

Graphic design is ubiquitous in people's daily lives. For graphic design, the most time-consuming task is laying out various components in the interface. Repetitive manual layout design will waste a lot of time for profe…

DisentanglementLayout DesignLayout Generation

Multiscale Structure-Guided Latent Diffusion for Multimodal MRI Translation

2026-03-13 · Jianqiang Lin, Zhiqiang Shen, Peng Cao, Jinzhu Yang 외 arxiv

Although diffusion models have achieved remarkable progress in multi-modal magnetic resonance imaging (MRI) translation tasks, existing methods still tend to suffer from anatomical inconsistencies or degraded texture det…

AirTrafficGen: Configurable Air Traffic Scenario Generation with Large Language Models

2025-08-04 · Dewi Sid William Gould, George De Ath, Ben Carvell, Nick Pepper arxiv

The manual design of scenarios for Air Traffic Control (ATC) training is a demanding and time-consuming bottleneck that limits the diversity of simulations available to controllers. To address this, we introduce a novel,…

Geometry-Consistent Neural Shape Representation with Implicit Displacement Fields

2021-06-09 · ICLR 2022 4 · Wang Yifan, Lukas Rahmann, Olga Sorkine-Hornung

We present implicit displacement fields, a novel representation for detailed 3D geometry. Inspired by a classic surface deformation technique, displacement mapping, our method represents a complex surface as a smooth bas…

3D geometryDisentanglementSurface Reconstruction

CalliMaster: Mastering Page-level Chinese Calligraphy via Layout-guided Spatial Planning

2026-03-12 · Tianshuo Xu, Tiantian Hong, Zhifei Chen, Fei Chao 외 arxiv

Page-level calligraphy synthesis requires balancing glyph precision with layout composition. Existing character models lack spatial context, while page-level methods often compromise brushwork detail. In this paper, we p…