paper-with-me

홈 › Papers

MUSE: Textual Attributes Guided Portrait Painting Generation

2020-11-09 · Xiaodan Hu, Pengfei Yu, Kevin Knight, Heng Ji, Bo Li, Honghui Shi

We propose a novel approach, MUSE, to illustrate textual attributes visually via portrait generation. MUSE takes a set of attributes written in text, in addition to facial features extracted from a photo of the subject as input. We propose 11 attribute types to represent inspirations from a subject's profile, emotion, story, and environment. We propose a novel stacked neural network architecture by extending an image-to-image generative model to accept textual attributes. Experiments show that our approach significantly outperforms several state-of-the-art methods without using textual attributes, with Inception Score score increased by 6% and Fr\'echet Inception Distance (FID) score decreased by 11%, respectively. We also propose a new attribute reconstruction metric to evaluate whether the generated portraits preserve the subject's attributes. Experiments show that our approach can accurately illustrate 78% textual attributes, which also help MUSE capture the subject in a more creative and expressive way.

📄 PDF Abstract BibTeX arXiv:2011.04761

Code (1)

xiaodanhu/MUSE 공식 구현 pytorch

Tasks

Attribute

Similar Papers 제목 키워드 기반

Recovering Faces from Portraits with Auxiliary Facial Attributes

2019-04-07 · Fatemeh Shiri, Xin Yu, Fatih Porikli, Richard Hartley 외

Recovering a photorealistic face from an artistic portrait is a challenging task since crucial facial details are often distorted or completely lost in artistic compositions. To handle this loss, we propose an Attribute-…

Attribute

How to Read Paintings: Semantic Art Understanding with Multi-Modal Retrieval

2018-10-23 · Noa Garcia, George Vogiatzis

Automatic art analysis has been mostly focused on classifying artworks into different artistic styles. However, understanding an artistic representation involves more complex processes, such as identifying the elements i…

Art AnalysisRetrieval

Structure-Guided Diffusion Models for High-Fidelity Portrait Shadow Removal

2025-07-07 · Wanchang Yu, Qing Zhang, Rongjia Zheng, Wei-Shi Zheng arxiv

We present a diffusion-based portrait shadow removal approach that can robustly produce high-fidelity results. Unlike previous methods, we cast shadow removal as diffusion-based inpainting. To this end, we first train a …

Shadow Removal

Inversion-Based Style Transfer with Diffusion Models

2022-11-23 · CVPR 2023 1 · Yuxin Zhang, Nisha Huang, Fan Tang, Haibin Huang 외

The artistic style within a painting is the means of expression, which includes not only the painting material, colors, and brushstrokes, but also the high-level attributes including semantic elements, object shapes, etc…

DenoisingImage GenerationStyle TransferText-to-Image Generation

MuseControlLite: Multifunctional Music Generation with Lightweight Conditioners

2025-06-23 · Fang-Duo Tsai, Shih-Lun Wu, Weijaw Lee, Sheng-Ping Yang 외

We propose MuseControlLite, a lightweight mechanism designed to fine-tune text-to-music generation models for precise conditioning using various time-varying musical attributes and reference audio signals. The key findin…

AttributeAudio inpaintingMusic GenerationText-to-Music Generation