paper-with-me

Papers

A Spatio-temporal Continuous Network for Stochastic 3D Human Motion Prediction

2025-08-03 · Hua Yu, Yaqing Hou, Xu Gui, Shanshan Feng, Dongsheng Zhou, Qiang Zhang arxiv

Stochastic Human Motion Prediction (HMP) has received increasing attention due to its wide applications. Despite the rapid progress in generative fields, existing methods often face challenges in learning continuous temporal dynamics and predicting stochastic motion sequences. They tend to overlook the flexibility inherent in complex human motions and are prone to mode collapse. To alleviate these issues, we propose a novel method called STCN, for stochastic and continuous human motion prediction, which consists of two stages. Specifically, in the first stage, we propose a spatio-temporal continuous network to generate smoother human motion sequences. In addition, the anchor set is innovatively introduced into the stochastic HMP task to prevent mode collapse, which refers to the potential human motion patterns. In the second stage, STCN endeavors to acquire the Gaussian mixture distribution (GMM) of observed motion sequences with the aid of the anchor set. It also focuses on the probability associated with each anchor, and employs the strategy of sampling multiple sequences from each anchor to alleviate intra-class differences in human motions. Experimental results on two widely-used datasets (Human3.6M and HumanEva-I) demonstrate that our model obtains competitive performance on both diversity and accuracy.

📄 PDF Abstract BibTeX arXiv:2508.01585

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

STDiff: Spatio-temporal Diffusion for Continuous Stochastic Video Prediction

2023-12-11 · Xi Ye, Guillaume-Alexandre Bilodeau

Predicting future frames of a video is challenging because it is difficult to learn the uncertainty of the underlying factors influencing their contents. In this paper, we propose a novel video prediction model, which ha…

PredictionVideo Prediction

MoChat: Joints-Grouped Spatio-Temporal Grounding LLM for Multi-Turn Motion Comprehension and Description

2024-10-15 · Jiawei Mo, Yixuan Chen, Rifen Lin, Yongkang Ni 외

Despite continuous advancements in deep learning for understanding human motion, existing models often struggle to accurately identify action timing and specific body parts, typically supporting only single-round interac…

Language ModelingLanguage ModellingLarge Language ModelMultimodal Large Language Model

Spatio-Temporal Branching for Motion Prediction using Motion Increments

2023-08-02 · Jiexin Wang, Yujie Zhou, Wenwen Qiang, Ying Ba 외

Human motion prediction (HMP) has emerged as a popular research topic due to its diverse applications, but it remains a challenging task due to the stochastic and aperiodic nature of future poses. Traditional methods rel…

Human motion predictionKnowledge Distillationmotion prediction

Bridge Frame and Event: Common Spatiotemporal Fusion for High-Dynamic Scene Optical Flow

2025-03-10 · CVPR 2025 1 · Hanyu Zhou, Haonan Wang, Haoyue Liu, Yuxing Duan 외

High-dynamic scene optical flow is a challenging task, which suffers spatial blur and temporal discontinuous motion due to large displacement in frame imaging, thus deteriorating the spatiotemporal feature of optical flo…

Optical Flow Estimation

DynTrace: Tracking Dynamic Object Evidence for 4D Spatio-Temporal Reasoning in MLLMs

2026-07-14 · Rongxin Gao, Yuzhi Huang, Dongxuan Liu, Chu Li 외 arxiv

4D spatio-temporal reasoning, jointly modeling 3D spatial structure and temporal evolution, is essential for understanding dynamic worlds and enabling embodied interaction. While current Multimodal Large Language Models …

Scene Understanding