paper-with-me

Papers Layout-to-Image Generation

“Layout-to-Image Generation” 태그가 달린 논문 50편 · 필터 해제

OccluRank: Controllable Occlusion-Aware Layout-to-Image Generation by Adding Just an Ordinal Rank

2026-08-21 · Wenyang Hong, Yuan Wang, Yanbin Hao, Lanqing Xue 외 arxiv

Layout-to-image generation enables explicit spatial control through bounding-box layouts, yet bounding boxes specify only instance locations and cannot represent their occlusion order. Existing methods may rely on additi…

Layout-to-Image Generation

Envisioning Beyond the Few: Disentangled Semantics and Primitives for Few-Shot Atypical Layout-to-Image Generation

2026-05-29 · Nan Bao, Yifan Zhao, Wenzhuang Wang, Jia Li arxiv

The layout-to-image (L2I) task enables fine-grained control over image generation via object categories and spatial layouts. However, existing L2I methods yield fragmented and distorted generations under few-shot atypica…

Layout-to-Image Generation

Visual Prototype Conditioned Focal Region Generation for UAV-Based Object Detection

2026-04-03 · Wenhao Li, Zimeng Wu, Yu Wu, Zehua Fu 외 arxiv

Unmanned aerial vehicle (UAV) based object detection is a critical but challenging task, when applied in dynamically changing scenarios with limited annotated training data. Layout-to-image generation approaches have pro…

Layout-to-Image GenerationObject Detection

EchoGen: Cycle-Consistent Learning for Unified Layout-Image Generation and Understanding

2026-03-18 · Kai Zou, Hongbo Liu, Dian Zheng, Jianxiong Gao 외 arxiv

In this work, we present EchoGen, a unified framework for layout-to-image generation and image grounding, capable of generating images with accurate layouts and high fidelity to text descriptions (e.g., spatial relations…

Layout-to-Image Generation

Laytrol: Preserving Pretrained Knowledge in Layout Control for Multimodal Diffusion Transformers

2025-11-11 · Sida Huang, Siqi Huang, Ping Luo, Hongyuan Zhang arxiv

With the development of diffusion models, enhancing spatial controllability in text-to-image generation has become a vital challenge. As a representative task for addressing this challenge, layout-to-image generation aim…

Layout-to-Image GenerationText-to-Image Generation

TerraGen: A Unified Multi-Task Layout Generation Framework for Remote Sensing Data Augmentation

2025-10-24 · Datao Tang, Hao Wang, Yudeng Xin, Hui Qiao 외 arxiv

Remote sensing vision tasks require extensive labeled data across multiple, interconnected domains. However, current generative data augmentation frameworks are task-isolated, i.e., each vision task requires training an …

Layout-to-Image GenerationData Augmentation

OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlaps

2025-09-23 · Bingnan Li, Chen-Yu Wang, Haiyang Xu, Xiang Zhang 외 arxiv

Despite steady progress in layout-to-image generation, current methods still struggle with layouts containing significant overlap between bounding boxes. We identify two primary challenges: (1) large overlapping regions …

Layout-to-Image Generation

InstanceAssemble: Layout-Aware Image Generation via Instance Assembling Attention

2025-09-20 · Qiang Xiang, Shuang Sun, Binglei Li, Dejia Song 외 arxiv

Diffusion models have demonstrated remarkable capabilities in generating high-quality images. Recent advancements in Layout-to-Image (L2I) generation have leveraged positional conditions and textual descriptions to facil…

Layout-to-Image Generation

LEARN: A Story-Driven Layout-to-Image Generation Framework for STEM Instruction

2025-08-15 · Maoquan Zhang, Bisser Raytchev, Xiujuan Sun arxiv

LEARN is a layout-aware diffusion framework designed to generate pedagogically aligned illustrations for STEM education. It leverages a curated BookCover dataset that provides narrative layouts and structured visual cues…

Layout-to-Image GenerationKnowledge Graphs

Psi-Sampler: Initial Particle Sampling for SMC-Based Inference-Time Reward Alignment in Score Models

2025-06-02 · Taehoon Yoon, Yunhong Min, Kyeongmin Yeo, Minhyuk Sung

We introduce $\Psi$-Sampler, an SMC-based framework incorporating pCNL-based initial particle sampling for effective inference-time reward alignment with a score-based generative model. Inference-time reward alignment wi…

DenoisingImage GenerationLayout-to-Image Generation

PlanGen: Towards Unified Layout Planning and Image Generation in Auto-Regressive Vision Language Models

2025-03-13 · Runze He, Bo Cheng, Yuhang Ma, Qingxiang Jia 외

In this paper, we propose a unified layout planning and image generation model, PlanGen, which can pre-plan spatial layout conditions before generating images. Unlike previous diffusion-based models that treat layout pla…

Image GenerationImage ManipulationLayout-to-Image Generation

ToLo: A Two-Stage, Training-Free Layout-To-Image Generation Framework For High-Overlap Layouts

2025-03-03 · Linhao Huang, Jing Yu

Recent training-free layout-to-image diffusion models have demonstrated remarkable performance in generating high-quality images with controllable layouts. These models follow a one-stage framework: Encouraging the model…

AttributeImage GenerationLayout-to-Image Generation

CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image Generation

2024-12-05 · HUI ZHANG, Dexiang Hong, Tingwei Gao, Yitong Wang 외

Diffusion models have been recognized for their ability to generate images that are not only visually appealing but also of high artistic quality. As a result, Layout-to-Image (L2I) generation has been proposed to levera…

Image GenerationLayout GenerationLayout-to-Image Generation

Boundary Attention Constrained Zero-Shot Layout-To-Image Generation

2024-11-15 · Huancheng Chen, Jingtao Li, Weiming Zhuang, Haris Vikalo 외

Recent text-to-image diffusion models excel at generating high-resolution images from text but struggle with precise control over spatial composition and object counting. To address these challenges, several studies deve…

Image GenerationLayout-to-Image GenerationObject Counting

HiCo: Hierarchical Controllable Diffusion Model for Layout-to-image Generation

2024-10-18 · Bo Cheng, Yuhang Ma, Liebucha Wu, Shanyuan Liu 외

The task of layout-to-image generation involves synthesizing images based on the captions of objects and their spatial positions. Existing methods still struggle in complex layout generation, where common bad cases inclu…

DisentanglementImage GenerationLayout GenerationLayout-to-Image Generation

Rethinking The Training And Evaluation of Rich-Context Layout-to-Image Generation

2024-09-07 · Jiaxin Cheng, Zixu Zhao, Tong He, Tianjun Xiao 외

Recent advancements in generative models have significantly enhanced their capacity for image generation, enabling a wide range of applications such as image editing, completion and video editing. A specialized area with…

Image GenerationLayout-to-Image GenerationVideo Editing

Training-free Composite Scene Generation for Layout-to-Image Synthesis

2024-07-18 · Jiaqi Liu, Tao Huang, Chang Xu

Recent breakthroughs in text-to-image diffusion models have significantly advanced the generation of high-fidelity, photo-realistic images from textual descriptions. Yet, these models often struggle with interpreting spa…

Image GenerationLayout-to-Image GenerationScene Generation

LTOS: Layout-controllable Text-Object Synthesis via Adaptive Cross-attention Fusions

2024-04-21 · Xiaoran Zhao, Tianhao Wu, Yu Lai, Zhiliang Tian 외

Controllable text-to-image generation synthesizes visual text and objects in images with certain conditions, which are frequently applied to emoji and poster generation. Visual text rendering and layout-to-image generati…

Image GenerationLayout-to-Image GenerationObjectText to Image Generation+1

ObjBlur: A Curriculum Learning Approach With Progressive Object-Level Blurring for Improved Layout-to-Image Generation

2024-04-11 · Stanislav Frolov, Brian B. Moser, Sebastian Palacio, Andreas Dengel

We present ObjBlur, a novel curriculum learning approach to improve layout-to-image generation models, where the task is to produce realistic images from layouts composed of boxes and labels. Our method is based on progr…

Image GenerationLayout-to-Image Generation

DivCon: Divide and Conquer for Progressive Text-to-Image Generation

2024-03-11 · Yuhao Jia, Wenhan Tan

Diffusion-driven text-to-image (T2I) generation has achieved remarkable advancements. To further improve T2I models' capability in numerical and spatial reasoning, the layout is employed as an intermedium to bridge large…

Image GenerationLayout-to-Image GenerationSpatial ReasoningText to Image Generation+1
1–20 / 50 다음 →