paper-with-me

Papers

Artistic Intelligence: A Diffusion-Based Framework for High-Fidelity Landscape Painting Synthesis

2024-07-24 · Wanggong Yang, Yifei Zhao

Generating high-fidelity landscape paintings remains a challenging task that requires precise control over both structure and style. In this paper, we present LPGen, a novel diffusion-based model specifically designed for landscape painting generation. LPGen introduces a decoupled cross-attention mechanism that independently processes structural and stylistic features, effectively mimicking the layered approach of traditional painting techniques. Additionally, LPGen proposes a structural controller, a multi-scale encoder designed to control the layout of landscape paintings, striking a balance between aesthetics and composition. Besides, the model is pre-trained on a curated dataset of high-resolution landscape images, categorized by distinct artistic styles, and then fine-tuned to ensure detailed and consistent output. Through extensive evaluations, LPGen demonstrates superior performance in producing paintings that are not only structurally accurate but also stylistically coherent, surpassing current state-of-the-art models. This work advances AI-generated art and offers new avenues for exploring the intersection of technology and traditional artistic practices. Our code, dataset, and model weights will be publicly available.

📄 PDF Abstract BibTeX arXiv:2407.17229

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderImage Generation

Methods 이 논문이 사용한 방법론

Latent Diffusion Model Diffusion models applied to latent spaces, which are normally built with (Variational) Autoencoders.
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

StyleProtect: Safeguarding Artistic Identity in Fine-tuned Diffusion Models

2025-09-17 · Qiuyu Tang, Joshua Krinsky, Aparna Bharati arxiv

The rapid advancement of generative models, particularly diffusion-based approaches, has inadvertently facilitated their potential for misuse. Such models enable malicious exploiters to replicate artistic styles that cap…

Towards Highly Realistic Artistic Style Transfer via Stable Diffusion with Step-aware and Layer-aware Prompt

2024-04-17 · Zhanjie Zhang, Quanwei Zhang, Huaizhong Lin, Wei Xing 외

Artistic style transfer aims to transfer the learned artistic style onto an arbitrary content image, generating artistic stylized images. Existing generative adversarial network-based methods fail to generate highly real…

Generative Adversarial NetworkStyle Transfer

VideoEraser: Concept Erasure in Text-to-Video Diffusion Models

2025-08-21 · Naen Xu, Jinghuai Zhang, Changjiang Li, Zhi Chen 외 arxiv

The rapid growth of text-to-video (T2V) diffusion models has raised concerns about privacy, copyright, and safety due to their potential misuse in generating harmful or misleading content. These models are often trained …

LineArt: A Knowledge-guided Training-free High-quality Appearance Transfer for Design Drawing with Diffusion Model

2024-12-16 · CVPR 2025 1 · Xi Wang, Hongzhen Li, Heng Fang, Yichen Peng 외

Image rendering from line drawings is vital in design and image generation technologies reduce costs, yet professional line drawings demand preserving complex details. Text prompts struggle with accuracy, and image trans…

Appearance TransferImage Generation

AvatarTex: High-Fidelity Facial Texture Reconstruction from Single-Image Stylized Avatars

2025-11-10 · Yuda Qiu, Zitong Xiao, Yiwei Zuo, Zisheng Ye 외 arxiv

We present AvatarTex, a high-fidelity facial texture reconstruction framework capable of generating both stylized and photorealistic textures from a single image. Existing methods struggle with stylized avatars due to th…