Layout-to-Image Generation
11개 벤치마크 · 논문 50편 · 이 태스크의 논문 보기 →
Benchmarks
COCO-Stuff 128x128
COCO-Stuff 256x256
COCO-Stuff 64x64
Visual Genome 128x128
LayoutBench-COCO - Size
Visual Genome 256x256
Visual Genome 64x64
LayoutBench
Most implemented
High-Resolution Image Synthesis with Latent Diffusion Models
Adding Conditional Control to Text-to-Image Diffusion Models
Image Synthesis From Reconfigurable Layout and Style
Image Generation from Scene Graphs
Papers
OccluRank: Controllable Occlusion-Aware Layout-to-Image Generation by Adding Just an Ordinal Rank
Layout-to-image generation enables explicit spatial control through bounding-box layouts, yet bounding boxes specify only instance locations and cannot represent their occlusion order. Existing methods may rely on additi…
Layout-to-Image GenerationEnvisioning Beyond the Few: Disentangled Semantics and Primitives for Few-Shot Atypical Layout-to-Image Generation
The layout-to-image (L2I) task enables fine-grained control over image generation via object categories and spatial layouts. However, existing L2I methods yield fragmented and distorted generations under few-shot atypica…
Layout-to-Image GenerationVisual Prototype Conditioned Focal Region Generation for UAV-Based Object Detection
Unmanned aerial vehicle (UAV) based object detection is a critical but challenging task, when applied in dynamically changing scenarios with limited annotated training data. Layout-to-image generation approaches have pro…
Layout-to-Image GenerationObject DetectionEchoGen: Cycle-Consistent Learning for Unified Layout-Image Generation and Understanding
In this work, we present EchoGen, a unified framework for layout-to-image generation and image grounding, capable of generating images with accurate layouts and high fidelity to text descriptions (e.g., spatial relations…
Layout-to-Image GenerationLaytrol: Preserving Pretrained Knowledge in Layout Control for Multimodal Diffusion Transformers
With the development of diffusion models, enhancing spatial controllability in text-to-image generation has become a vital challenge. As a representative task for addressing this challenge, layout-to-image generation aim…
Layout-to-Image GenerationText-to-Image GenerationTerraGen: A Unified Multi-Task Layout Generation Framework for Remote Sensing Data Augmentation
Remote sensing vision tasks require extensive labeled data across multiple, interconnected domains. However, current generative data augmentation frameworks are task-isolated, i.e., each vision task requires training an …
Layout-to-Image GenerationData Augmentation