paper-with-me

홈 › Papers

MultiAct: Long-Term 3D Human Motion Generation from Multiple Action Labels

2022-12-12 · Taeryung Lee, Gyeongsik Moon, Kyoung Mu Lee

We tackle the problem of generating long-term 3D human motion from multiple action labels. Two main previous approaches, such as action- and motion-conditioned methods, have limitations to solve this problem. The action-conditioned methods generate a sequence of motion from a single action. Hence, it cannot generate long-term motions composed of multiple actions and transitions between actions. Meanwhile, the motion-conditioned methods generate future motions from initial motion. The generated future motions only depend on the past, so they are not controllable by the user's desired actions. We present MultiAct, the first framework to generate long-term 3D human motion from multiple action labels. MultiAct takes account of both action and motion conditions with a unified recurrent generation system. It repetitively takes the previous motion and action label; then, it generates a smooth transition and the motion of the given action. As a result, MultiAct produces realistic long-term motion controlled by the given sequence of multiple action labels. Codes are available here at https://github.com/TaeryungLee/MultiAct_RELEASE.

📄 PDF Abstract BibTeX arXiv:2212.05897

Code (1)

TaeryungLee/MultiAct_RELEASE 공식 구현 pytorch

Tasks

Motion Generation

Similar Papers 제목 키워드 기반

MultiAct: Text-to-Motion Generation from Composite Text via Tailored Attention Guidance

2026-05-29 · Nathan Sala, Ofir Abramovich, Ariel Shamir, Daniel Cohen-Or 외 arxiv

Text-to-motion generation has progressed rapidly in recent years, offering an expressive interface for animation and human-computer interaction. However, current models remain brittle when handling prompts that describe …

Motion Synthesis

MultiActor-Audiobook: Zero-Shot Audiobook Generation with Faces and Voices of Multiple Speakers

2025-05-19 · Kyeongman Park, Seongho Joo, Kyomin Jung

We introduce MultiActor-Audiobook, a zero-shot approach for generating audiobooks that automatically produces consistent, expressive, and speaker-appropriate prosody, including intonation and emotion. Previous audiobook …

Sentence

GlocalNet: Class-aware Long-term Human Motion Synthesis

2020-12-19 · Neeraj Battan, Yudhik Agrawal, Veeravalli Saisooryarao, Aman Goel 외

Synthesis of long-term human motion skeleton sequences is essential to aid human-centric video generation with potential applications in Augmented Reality, 3D character animations, pedestrian trajectory prediction, etc. …

Motion SynthesisPedestrian Trajectory PredictionTrajectory PredictionVideo Generation

T2LM: Long-Term 3D Human Motion Generation from Multiple Sentences

2024-06-02 · Taeryung Lee, Fabien Baradel, Thomas Lucas, Kyoung Mu Lee 외

In this paper, we address the challenging problem of long-term 3D human motion generation. Specifically, we aim to generate a long sequence of smoothly connected actions from a stream of multiple sentences (i.e., paragra…

Action GenerationDecoderMotion Generation

Cross-Conditioned Recurrent Networks for Long-Term Synthesis of Inter-Person Human Motion Interactions

2020-05-14 · IEEE Winter Conference on Applications of Computer Vision (WACV) 2020 5 · Jogendra Nath Kundu, Himanshu Buckchash, Priyanka Mandikal, Anirudh Jamkhandi 외

Modeling dynamics of human motion is one of the most challenging sequence modeling problem, with diverse applications in animation industry, human-robot interaction, motion-based surveillance, etc. Available attempts to …

DecoderMotion Generationmotion predictionMotion Synthesis+1