paper-with-me

Papers Motion Generation

“Motion Generation” 태그가 달린 논문 446편 · 필터 해제

SnapMoGen: Human Motion Generation from Expressive Texts

2025-07-12 · Chuan Guo, Inwoo Hwang, Jian Wang, Bing Zhou

Text-to-motion generation has experienced remarkable progress in recent years. However, current approaches remain limited to synthesizing motion from short or general text prompts, primarily due to dataset constraints. T…

Motion Generation

Go to Zero: Towards Zero-shot Motion Generation with Million-scale Data

2025-07-09 · Ke Fan, Shunlin Lu, Minyue Dai, Runyi Yu 외

Generating diverse and natural human motion sequences based on textual descriptions constitutes a fundamental and challenging research area within the domains of computer vision, graphics, and robotics. Despite significa…

Motion GenerationZero-shot Generalization

Motion Generation: A Survey of Generative Approaches and Benchmarks

2025-07-07 · Aliasghar Khani, Arianna Rampini, Bruno Roy, Larasika Nadela 외

Motion generation, the task of synthesizing realistic motion sequences from various conditioning inputs, has become a central problem in computer vision, computer graphics, and robotics, with applications ranging from an…

Motion GenerationSurvey

A Unified Transformer-Based Framework with Pretraining For Whole Body Grasping Motion Generation

2025-07-01 · Edward Effendy, Kuan-Wei Tseng, Rei Kawakami

Accepted in the ICIP 2025 We present a novel transformer-based framework for whole-body grasping that addresses both pose generation and motion infilling, enabling realistic and stable object interactions. Our pipeline c…

Grasp GenerationMotion Generation

PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis

2025-06-22 · Chuhao Jin, Haosen Li, Bingzi Zhang, Che Liu 외

Recent advances in large language models (LLMs) have enabled breakthroughs in many multimodal generation tasks, but a significant performance gap still exists in text-to-motion generation, where LLM-based methods lag far…

DiversityMotion GenerationMotion Synthesismultimodal generation

Human-Centered Editable Speech-to-Sign-Language Generation via Streaming Conformer-Transformer and Resampling Hook

2025-06-17 · Yingchao Li

Existing end-to-end sign-language animation systems suffer from low naturalness, limited facial/body expressivity, and no user control. We propose a human-centered, real-time speech-to-sign animation framework that integ…

Motion GenerationText Generation

RL from Physical Feedback: Aligning Large Motion Models with Humanoid Control

2025-06-15 · Junpeng Yue, Zepeng Wang, Yuxuan Wang, Weishuai Zeng 외

This paper focuses on a critical challenge in robotics: translating text-driven human motions into executable actions for humanoid robots, enabling efficient and cost-effective learning of new behaviors. While existing t…

Humanoid ControlMotion GenerationSemantic correspondence

Motion-R1: Chain-of-Thought Reasoning and Reinforcement Learning for Human Motion Generation

2025-06-12 · Runqi Ouyang, Haoyun Li, Zhenyuan Zhang, XiaoFeng Wang 외

Recent advances in large language models, especially in natural language understanding and reasoning, have opened new possibilities for text-to-motion generation. Although existing approaches have made notable progress i…

Language ModelingLanguage ModellingLogical ReasoningMotion Generation+2

PhysiInter: Integrating Physical Mapping for High-Fidelity Human Interaction Generation

2025-06-09 · Wei Yao, Yunlian Sun, Chang Liu, Hongwen Zhang 외

Driven by advancements in motion capture and generative artificial intelligence, leveraging large-scale MoCap datasets to train generative models for synthesizing diverse, realistic human motions has become a promising r…

Motion Generationvalid

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios

2025-06-03 · Lingwei Dang, Ruizhi Shao, Hongwen Zhang, Wei Min 외

Hand-Object Interaction (HOI) generation has significant application potential. However, current 3D HOI motion generation approaches heavily rely on predefined 3D object models and lab-captured motion data, limiting gene…

Motion GenerationVideo Generation

UniConFlow: A Unified Constrained Generalization Framework for Certified Motion Planning with Flow Matching Models

2025-06-03 · Zewen Yang, Xiaobing Dai, Dian Yu, Qianru Li 외

Generative models have become increasingly powerful tools for robot motion generation, enabling flexible and multimodal trajectory generation across various tasks. Yet, most existing approaches remain limited in handling…

Collision AvoidanceMotion GenerationMotion Planning

Captivity-Escape Games as a Means for Safety in Online Motion Generation

2025-06-02 · Christopher Bohn, Manuel Hess, Sören Hohmann

This paper presents a method that addresses the conservatism, computational effort, and limited numerical accuracy of existing frameworks and methods that ensure safety in online model-based motion generation, commonly r…

Motion GenerationMotion Planning

EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models

2025-06-02 · Andy Bonnetto, Haozhe Qi, Franklin Leong, Matea Tashkovska 외

Understanding behavior requires datasets that capture humans while carrying out complex tasks. The kitchen is an excellent environment for assessing human motor and cognitive function, as many complex actions are natural…

Action RecognitionAction SegmentationMotion Generation

Semantics-Aware Human Motion Generation from Audio Instructions

2025-05-29 · Zi-An Wang, Shihao Zou, Shiyao Yu, Mingyuan Zhang 외

Recent advances in interactive technologies have highlighted the prominence of audio signals for semantic encoding. This paper explores a new task, where audio signals are used as conditioning inputs to generate motions …

Motion Generation

Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation

2025-05-29 · CVPR 2025 1 · Hao Li, Ju Dai, Xin Zhao, Feng Zhou 외

In 3D speech-driven facial animation generation, existing methods commonly employ pre-trained self-supervised audio models as encoders. However, due to the prevalence of phonetically similar syllables with distinct lip s…

Motion Generation

MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation

2025-05-29 · Siyuan Wang, Jiawei Liu, Wei Wang, Yeying Jin 외

Co-Speech Gesture Video Generation aims to generate vivid speech videos from audio-driven still images, which is challenging due to the diversity of different parts of the body in terms of amplitude of motion, audio rele…

Motion GenerationVideo Generation

From Motion to Behavior: Hierarchical Modeling of Humanoid Generative Behavior Control

2025-05-28 · Jusheng Zhang, Jinzhou Tang, Sidi Liu, Mingyan Li 외

Human motion generative modeling or synthesis aims to characterize complicated human motions of daily activities in diverse real-world environments. However, current research predominantly focuses on either low-level, sh…

Motion GenerationMotion PlanningTask and Motion Planning

IKMo: Image-Keyframed Motion Generation with Trajectory-Pose Conditioned Motion Diffusion Model

2025-05-27 · Yang Zhao, Yan Zhang, Xubo Yang

Existing human motion generation methods with trajectory and pose inputs operate global processing on both modalities, leading to suboptimal outputs. In this paper, we propose IKMo, an image-keyframed motion generation m…

Motion Generation

Absolute Coordinates Make Motion Generation Easy

2025-05-26 · Zichong Meng, Zeyu Han, Xiaogang Peng, Yiming Xie 외

State-of-the-art text-to-motion generation models rely on the kinematic-aware, local-relative motion representation popularized by HumanML3D, which encodes motion relative to the pelvis and to the previous frame with bui…

Motion Generation

From Single Images to Motion Policies via Video-Generation Environment Representations

2025-05-25 · Weiming Zhi, Ziyong Ma, Tianyi Zhang, Matthew Johnson-Roberson

Autonomous robots typically need to construct representations of their surroundings and adapt their motions to the geometry of their environment. Here, we tackle the problem of constructing a policy model for collision-f…

Depth EstimationMonocular Depth EstimationMotion GenerationVideo Generation
1–20 / 446 다음 →