paper-with-me

홈 › Papers

Learning to Animate Images from A Few Videos to Portray Delicate Human Actions

2025-03-01 · Haoxin Li, Yingchen Yu, Qilong Wu, Hanwang Zhang, Boyang Li, Song Bai

Despite recent progress, video generative models still struggle to animate human actions from static images, particularly when handling uncommon actions whose training data are limited. In this paper, we investigate the task of learning to animate human actions from a small number of videos -- 16 or fewer -- which is highly valuable in real-world applications like video and movie production. Few-shot learning of generalizable motion patterns while ensuring smooth transitions from the initial reference image is exceedingly challenging. We propose FLASH (Few-shot Learning to Animate and Steer Humans), which improves motion generalization by aligning motion features and inter-frame correspondence relations between videos that share the same motion but have different appearances. This approach minimizes overfitting to visual appearances in the limited training data and enhances the generalization of learned motion patterns. Additionally, FLASH extends the decoder with additional layers to compensate lost details in the latent space, fostering smooth transitions from the initial reference image. Experiments demonstrate that FLASH effectively animates images with unseen human or scene appearances into specified actions while maintaining smooth transitions from the reference image.

📄 PDF Abstract BibTeX arXiv:2503.00276

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderFew-Shot Learning

Similar Papers 제목 키워드 기반

Can Pose Transfer Models Generate Realistic Human Motion?

2025-01-26 · Vaclav Knapp, Matyas Bohacek

Recent pose-transfer methods aim to generate temporally consistent and fully controllable videos of human action where the motion from a reference video is reenacted by a new identity. We evaluate three state-of-the-art …

Pose Transfer

CAST: Character labeling in Animation using Self-supervision by Tracking

2022-01-19 · Oron Nir, Gal Rapoport, Ariel Shamir

Cartoons and animation domain videos have very different characteristics compared to real-life images and videos. In addition, this domain carries a large variability in styles. Current computer vision and deep-learning …

Multi-Object TrackingObject TrackingRepresentation Learning

Animate-X++: Universal Character Image Animation with Dynamic Backgrounds

2025-08-13 · Shuai Tan, Biao Gong, Zhuoxin Liu, Yan Wang 외 arxiv

Character image animation, which generates high-quality videos from a reference image and target pose sequence, has seen significant progress in recent years. However, most existing methods only apply to human figures, w…

Synthetic Human Memories: AI-Edited Images and Videos Can Implant False Memories and Distort Recollection

2024-09-13 · Pat Pataranutaporn, Chayapatr Archiwaranguprok, Samantha W. T. Chan, Elizabeth Loftus 외

AI is increasingly used to enhance images and videos, both intentionally and unintentionally. As AI editing tools become more integrated into smartphones, users can modify or animate photos into realistic videos. This st…

AnimateScene: Camera-controllable Animation in Any Scene

2025-08-08 · Qingyang Liu, Bingjie Gao, Weiheng Huang, Jun Zhang 외 arxiv

Recent advances in 3D scene reconstruction and 4D human animation have broadened adoption, but integrating the two remains difficult. Key challenges include placing humans at plausible locations and scales without interp…