Papers Layout Generation
“Layout Generation” 태그가 달린 논문 148편 · 필터 해제
Multi-Agent Synergy-Driven Iterative Visual Narrative Synthesis
Automated generation of high-quality media presentations is challenging, requiring robust content extraction, narrative planning, visual design, and overall quality optimization. Existing methods often produce presentati…
Layout GenerationReLayout: Integrating Relation Reasoning for Content-aware Layout Generation with Multi-modal Large Language Models
Content-aware layout aims to arrange design elements appropriately on a given canvas to convey information effectively. Recently, the trend for this task has been to leverage large language models (LLMs) to generate layo…
Layout GenerationEmbodiedGen: Towards a Generative 3D World Engine for Embodied Intelligence
Constructing a physically realistic and accurately scaled simulated 3D world is crucial for the training and evaluation of embodied intelligence tasks. The diversity, realism, low cost accessibility and affordability of …
Image to 3DLayout GenerationScene GenerationText to 3D+1LLM-driven Indoor Scene Layout Generation via Scaled Human-aligned Data Synthesis and Multi-Stage Preference Optimization
Automatic indoor layout generation has attracted increasing attention due to its potential in interior design, virtual environment construction, and embodied AI. Existing methods fall into two categories: prompt-driven a…
Layout GenerationRobot NavigationDirect Numerical Layout Generation for 3D Indoor Scene Synthesis via Spatial Reasoning
Realistic 3D indoor scene synthesis is vital for embodied AI and digital content creation. It can be naturally divided into two subtasks: object generation and layout generation. While recent generative models have signi…
In-Context LearningIndoor Scene SynthesisLayout GenerationObject+1Aggregated Structural Representation with Large Language Models for Human-Centric Layout Generation
Time consumption and the complexity of manual layout design make automated layout generation a critical task, especially for multiple applications across different mobile devices. Existing graph-based layout generation a…
Layout DesignLayout GenerationTextDiffuser-RL: Efficient and Robust Text Layout Optimization for High-Fidelity Text-to-Image Synthesis
Text-embedded image generation plays a critical role in industries such as graphic design, advertising, and digital content creation. Text-to-Image generation methods leveraging diffusion models, such as TextDiffuser-2, …
CPUGPUImage GenerationLayout Generation+4Lay-Your-Scene: Natural Scene Layout Generation with Diffusion Transformers
We present Lay-Your-Scene (shorthand LayouSyn), a novel text-to-layout generation pipeline for natural scenes. Prior scene layout generation methods are either closed-vocabulary or use proprietary large language models f…
Image GenerationLayout GenerationPosterO: Structuring Layout Trees to Enable Language Models in Generalized Content-Aware Layout Generation
In poster design, content-aware layout generation is crucial for automatically arranging visual-textual elements on the given image. With limited training data, existing work focused on image-centric enhancement. However…
In-Context LearningLayout GenerationGenerating Animated Layouts as Structured Text Representations
Despite the remarkable progress in text-to-video models, achieving precise control over text elements and animated graphics remains a significant challenge, especially in applications such as video advertisements. To add…
Layout GenerationLayoutCoT: Unleashing the Deep Reasoning Potential of Large Language Models for Layout Generation
Conditional layout generation aims to automatically generate visually appealing and semantically coherent layouts from user-defined constraints. While recent methods based on generative models have shown promising result…
In-Context LearningLayout GenerationRAGRetrieval+1Relation-Rich Visual Document Generator for Visual Information Extraction
Despite advances in Large Language Models (LLMs) and Multimodal LLMs (MLLMs) for visual document understanding (VDU), visual information extraction (VIE) from relation-rich documents remains challenging due to the layout…
Diversitydocument understandingLayout GenerationOptical Character Recognition+2Hierarchical and Step-Layer-Wise Tuning of Attention Specialty for Multi-Instance Synthesis in Diffusion Transformers
Text-to-image (T2I) generation models often struggle with multi-instance synthesis (MIS), where they must accurately depict multiple distinct instances in a single image based on complex prompts detailing individual feat…
AttributeLayout GenerationDecoupled Diffusion Sparks Adaptive Scene Generation
Controllable scene generation could reduce the cost of diverse data collection substantially for autonomous driving. Prior works formulate the traffic layout generation as predictive progress, either by denoising entire …
Autonomous DrivingData AugmentationDenoisingLayout Generation+1Computer-Aided Layout Generation for Building Design: A Review
Generating realistic building layouts for automatic building design has been studied in both the computer vision and architecture domains. Traditional approaches from the architecture domain, which are based on optimizat…
Layout DesignLayout GenerationEfficient Multi-Instance Generation with Janus-Pro-Dirven Prompt Parsing
Recent advances in text-guided diffusion models have revolutionized conditional image generation, yet they struggle to synthesize complex scenes with multiple objects due to imprecise spatial grounding and limited scalab…
Conditional Image GenerationImage GenerationLayout GenerationRelTriple: Learning Plausible Indoor Layouts by Integrating Relationship Triples into the Diffusion Process
The generation of indoor furniture layouts has significant applications in augmented reality, smart homes, and architectural design. Successful furniture arrangement requires proper physical relationships (e.g., collisio…
Collision AvoidanceLayout GenerationObjectBADGR: Bundle Adjustment Diffusion Conditioned by GRadients for Wide-Baseline Floor Plan Reconstruction
Reconstructing precise camera poses and floor plan layouts from wide-baseline RGB panoramas is a difficult and unsolved problem. We introduce BADGR, a novel diffusion model that jointly performs reconstruction and bundle…
DenoisingLayout GenerationMulti-View 3D ReconstructionMulti-view Floor Layout Reconstruction (N-view)+2MetaSpatial: Reinforcing 3D Spatial Reasoning in VLMs for the Metaverse
We present MetaSpatial, the first reinforcement learning (RL)-based framework designed to enhance 3D spatial reasoning in vision-language models (VLMs), enabling real-time 3D scene generation without the need for hard-co…
Layout GenerationReinforcement Learning (RL)Scene GenerationSpatial ReasoningHSM: Hierarchical Scene Motifs for Multi-Scale Indoor Scene Generation
Despite advances in indoor 3D scene layout generation, synthesizing scenes with dense object arrangements remains challenging. Existing methods primarily focus on large furniture while neglecting smaller objects, resulti…
Layout GenerationScene Generation