paper-with-me

Papers

Bidirectionally Deformable Motion Modulation For Video-based Human Pose Transfer

2023-07-15 · ICCV 2023 1 · Wing-Yin Yu, Lai-Man Po, Ray C. C. Cheung, Yuzhi Zhao, Yu Xue, Kun Li

Video-based human pose transfer is a video-to-video generation task that animates a plain source human image based on a series of target human poses. Considering the difficulties in transferring highly structural patterns on the garments and discontinuous poses, existing methods often generate unsatisfactory results such as distorted textures and flickering artifacts. To address these issues, we propose a novel Deformable Motion Modulation (DMM) that utilizes geometric kernel offset with adaptive weight modulation to simultaneously perform feature alignment and style transfer. Different from normal style modulation used in style transfer, the proposed modulation mechanism adaptively reconstructs smoothed frames from style codes according to the object shape through an irregular receptive field of view. To enhance the spatio-temporal consistency, we leverage bidirectional propagation to extract the hidden motion information from a warped image sequence generated by noisy poses. The proposed feature propagation significantly enhances the motion prediction ability by forward and backward propagation. Both quantitative and qualitative experimental results demonstrate superiority over the state-of-the-arts in terms of image fidelity and visual continuity. The source code is publicly available at github.com/rocketappslab/bdmm.

📄 PDF Abstract BibTeX arXiv:2307.07754

Code (1)

rocketappslab/bdmm 공식 구현 pytorch

Tasks

motion predictionPose TransferStyle TransferVideo Generation

Similar Papers 제목 키워드 기반

Temporal Modulation Network for Controllable Space-Time Video Super-Resolution

2021-04-21 · CVPR 2021 1 · Gang Xu, Jun Xu, Zhen Li, Liang Wang 외

Space-time video super-resolution (STVSR) aims to increase the spatial and temporal resolutions of low-resolution and low-frame-rate videos. Recently, deformable convolution based methods have achieved promising STVSR pe…

Space-time Video Super-resolutionSuper-ResolutionVideo Super-Resolution

BridgeSplat: Bidirectionally Coupled CT and Non-Rigid Gaussian Splatting for Deformable Intraoperative Surgical Navigation

2025-09-23 · Maximilian Fehrentz, Alexander Winkler, Thomas Heiliger, Nazim Haouchine 외 arxiv

We introduce BridgeSplat, a novel approach for deformable surgical navigation that couples intraoperative 3D reconstruction with preoperative CT data to bridge the gap between surgical video and volumetric patient data. …

3D Reconstruction

MoCapDeform: Monocular 3D Human Motion Capture in Deformable Scenes

2022-08-17 · Zhi Li, Soshi Shimada, Bernt Schiele, Christian Theobalt 외

3D human motion capture from monocular RGB images respecting interactions of a subject with complex and possibly deformable environments is a very challenging, ill-posed and under-explored problem. Existing methods addre…

3D Human Pose EstimationPose Estimation

Structure From Tracking: Distilling Structure-Preserving Motion for Video Generation

2025-12-12 · Yang Fei, George Stoica, Jingyuan Liu, Qifeng Chen 외 arxiv

Reality is a dance between rigid constraints and deformable structures. For video models, that means generating motion that preserves fidelity as well as structure. Despite progress in diffusion models, producing realist…

Video Generation

Sign Language Recognition via Deformable 3D Convolutions and Modulated Graph Convolutional Networks

2023-06-07 · IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) 2023 6 · Katerina Papadimitriou, Gerasimos Potamianos

Automatic sign language recognition (SLR) remains challenging, especially when employing RGB video alone (i.e., with no depth or special glove-based input) and under a signer-independent (SI) framework, due to inter-pers…

graph constructionSign Language Recognition