paper-with-me

홈 › Papers

Cross Attention Based Style Distribution for Controllable Person Image Synthesis

2022-08-01 · Xinyue Zhou, Mingyu Yin, Xinyuan Chen, Li Sun, Changxin Gao, Qingli Li

Controllable person image synthesis task enables a wide range of applications through explicit control over body pose and appearance. In this paper, we propose a cross attention based style distribution module that computes between the source semantic styles and target pose for pose transfer. The module intentionally selects the style represented by each semantic and distributes them according to the target pose. The attention matrix in cross attention expresses the dynamic similarities between the target pose and the source styles for all semantics. Therefore, it can be utilized to route the color and texture from the source image, and is further constrained by the target parsing map to achieve a clearer objective. At the same time, to encode the source appearance accurately, the self attention among different semantic styles is also added. The effectiveness of our model is validated quantitatively and qualitatively on pose transfer and virtual try-on tasks.

📄 PDF Abstract BibTeX arXiv:2208.00712

Code (1)

xyzhouo/casd 공식 구현 pytorch

Tasks

Image GenerationPose TransferVirtual Try-on

Similar Papers 제목 키워드 기반

Voicing Personas: Rewriting Persona Descriptions into Style Prompts for Controllable Text-to-Speech

2025-05-21 · Yejin Lee, Jaehoon Kang, Kyuhong Shim

In this paper, we propose a novel framework to control voice style in prompt-based, controllable text-to-speech systems by leveraging textual personas as voice style prompts. We present two persona rewriting strategies t…

text-to-speechText to Speech

StyleFusion360: View-Consistent Head Stylization via Adaptive Style Modulation

2025-11-27 · Furkan Guzelant, Arda Goktogan, Tarık Kaya, Aysegul Dundar arxiv

3D head stylization enables expressive reimagining of human faces for creative visual experiences in digital media. Existing 3D-aware methods often require computationally intensive optimization or per-style fine-tuning,…

StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking Heads

2024-09-14 · Suzhen Wang, Yifeng Ma, Yu Ding, Zhipeng Hu 외

Individuals have unique facial expression and head pose styles that reflect their personalized speaking styles. Existing one-shot talking head methods cannot capture such personalized characteristics and therefore fail t…

Face GenerationTalking Face Generation

Training-Free Multi-Style Fusion Through Reference-Based Adaptive Modulation

2025-09-23 · Xu Liu, Yibo Lu, Xinxian Wang, Xinyu Wu arxiv

We propose Adaptive Multi-Style Fusion (AMSF), a reference-based training-free framework that enables controllable fusion of multiple reference styles in diffusion models. Most of the existing reference-based methods are…

Content-style disentangled representation for controllable artistic image stylization and generation

2024-12-19 · Ma Zhuoqi, Zhang Yixuan, You Zejun, Tian Long 외

Controllable artistic image stylization and generation aims to render the content provided by text or image with the learned artistic style, where content and style decoupling is the key to achieve satisfactory results. …

DisentanglementImage Stylization