paper-with-me

홈 › Papers

Deep Spatial Transformation for Pose-Guided Person Image Generation and Animation

2020-08-27 · Yurui Ren, Ge Li, Shan Liu, Thomas H. Li

Pose-guided person image generation and animation aim to transform a source person image to target poses. These tasks require spatial manipulation of source data. However, Convolutional Neural Networks are limited by the lack of ability to spatially transform the inputs. In this paper, we propose a differentiable global-flow local-attention framework to reassemble the inputs at the feature level. This framework first estimates global flow fields between sources and targets. Then, corresponding local source feature patches are sampled with content-aware local attention coefficients. We show that our framework can spatially transform the inputs in an efficient manner. Meanwhile, we further model the temporal consistency for the person image animation task to generate coherent videos. The experiment results of both image generation and animation tasks demonstrate the superiority of our model. Besides, additional results of novel view synthesis and face image animation show that our model is applicable to other tasks requiring spatial transformation. The source code of our project is available at https://github.com/RenYurui/Global-Flow-Local-Attention.

📄 PDF Abstract BibTeX arXiv:2008.12606

Code (1)

RenYurui/Global-Flow-Local-Attention 공식 구현 pytorch

Tasks

Image AnimationImage GenerationNovel View Synthesis

Similar Papers 제목 키워드 기반

Deep Image Spatial Transformation for Person Image Generation

2020-03-02 · CVPR 2020 6 · Yurui Ren, Xiaoming Yu, Junming Chen, Thomas H. Li 외

Pose-guided person image generation is to transform a source person image to a target pose. This task requires spatial manipulations of source data. However, Convolutional Neural Networks are limited by the lack of abili…

Image Generation

Combining Attention with Flow for Person Image Synthesis

2021-08-04 · Yurui Ren, Yubo Wu, Thomas H. Li, Shan Liu 외

Pose-guided person image synthesis aims to synthesize person images by transforming reference images into target poses. In this paper, we observe that the commonly used spatial transformation blocks have complementary ad…

Image Generation

Soft-Gated Warping-GAN for Pose-Guided Person Image Synthesis

2018-10-27 · NeurIPS 2018 12 · Haoye Dong, Xiaodan Liang, Ke Gong, Hanjiang Lai 외

Despite remarkable advances in image synthesis research, existing works often fail in manipulating images under the context of large geometric transformations. Synthesizing person images conditioned on arbitrary poses is…

Generative Adversarial NetworkImage Generation

Structure-aware Person Image Generation with Pose Decomposition and Semantic Correlation

2021-02-05 · Jilin Tang, Yi Yuan, Tianjia Shao, Yong liu 외

In this paper we tackle the problem of pose guided person image generation, which aims to transfer a person image from the source pose to a novel target pose while maintaining the source appearance. Given the inefficienc…

Image Generation

Talk2Move: Reinforcement Learning for Text-Instructed Object-Level Geometric Transformation in Scenes

2026-01-05 · Jing Tan, Zhaoyang Zhang, Yantao Shen, Jiarui Cai 외 arxiv

We introduce Talk2Move, a reinforcement learning (RL) based diffusion framework for text-instructed spatial transformation of objects within scenes. Spatially manipulating objects in a scene through natural language pose…

Reinforcement Learningmultimodal generation