paper-with-me

Papers

OmniPSD: Layered PSD Generation with Diffusion Transformer

2025-12-10 · Cheng Liu, Yiren Song, Haofan Wang, Mike Zheng Shou arxiv

Recent advances in diffusion models have greatly improved image generation and editing, yet generating or reconstructing layered PSD files with transparent alpha channels remains highly challenging. We propose OmniPSD, a unified diffusion framework built upon the Flux ecosystem that enables both text-to-PSD generation and image-to-PSD decomposition through in-context learning. For text-to-PSD generation, OmniPSD arranges multiple target layers spatially into a single canvas and learns their compositional relationships through spatial attention, producing semantically coherent and hierarchically structured layers. For image-to-PSD decomposition, it performs iterative in-context editing, progressively extracting and erasing textual and foreground components to reconstruct editable PSD layers from a single flattened image. An RGBA-VAE is employed as an auxiliary representation module to preserve transparency without affecting structure learning. Extensive experiments on our new RGBA-layered dataset demonstrate that OmniPSD achieves high-fidelity generation, structural consistency, and transparency awareness, offering a new paradigm for layered design generation and decomposition with diffusion transformers.

📄 PDF Abstract BibTeX arXiv:2512.09247

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

LaDe: Unified Multi-Layered Graphic Media Generation and Decomposition

2026-03-18 · Vlad-Constantin Lungu-Stan, Ionut Mironica, Mariana-Iuliana Georgescu arxiv

Media design layer generation enables the creation of fully editable, layered design documents such as posters, flyers, and logos using only natural language prompts. Existing methods either restrict outputs to a fixed n…

Text-to-Image Generation

LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer

2025-02-03 · Yiren Song, Danze Chen, Mike Zheng Shou

Generating cognitive-aligned layered SVGs remains challenging due to existing methods' tendencies toward either oversimplified single-layer outputs or optimization-induced shape redundancies. We propose LayerTracer, a di…

Vera: A Layered Diffusion Model for Content-Preserving Video Editing

2026-06-22 · Hongkai Zheng, Ta-Ying Cheng, Benjamin Klein, Yisong Yue 외 arxiv

Video diffusion models have enabled remarkable progress in video generation and editing. However, content preservation remains a core challenge: existing methods regenerate every pixel and often alter elements that shoul…

Video Generation

MagicQuillV2: Precise and Interactive Image Editing with Layered Visual Cues

2025-12-02 · Zichen Liu, Yue Yu, Hao Ouyang, Qiuyu Wang 외 arxiv

We propose MagicQuill V2, a novel system that introduces a \textbf{layered composition} paradigm to generative image editing, bridging the gap between the semantic power of diffusion models and the granular control of tr…

Image Editing

Text2Layer: Layered Image Generation using Latent Diffusion Model

2023-07-19 · Xinyang Zhang, Wentian Zhao, Xin Lu, Jeff Chien

Layer compositing is one of the most popular image editing workflows among both amateurs and professionals. Motivated by the success of diffusion models, we explore layer compositing from a layered image generation persp…

Image GenerationImage SegmentationmodelSemantic Segmentation