paper-with-me

Papers

HardMo: A Large-Scale Hardcase Dataset for Motion Capture

2024-01-01 · CVPR 2024 1 · Jiaqi Liao, Chuanchen Luo, Yinuo Du, Yuxi Wang, XuCheng Yin, Man Zhang, Zhaoxiang Zhang, Junran Peng

Recent years have witnessed rapid progress in monocular human mesh recovery. Despite their impressive performance on public benchmarks existing methods are vulnerable to unusual poses which prevents them from deploying to challenging scenarios such as dance and martial arts. This issue is mainly attributed to the domain gap induced by the data scarcity in relevant cases. Most existing datasets are captured in constrained scenarios and lack samples of such complex movements. For this reason we propose a data collection pipeline comprising automatic crawling precise annotation and hardcase mining. Based on this pipeline we establish a large dataset in a short time. The dataset named HardMo contains 7M images along with precise annotations covering 15 categories of dance and 14 categories of martial arts. Empirically we find that the prediction failure in dance and martial arts is mainly characterized by the misalignment of hand-wrist and foot-ankle. To dig deeper into the two hardcases we leverage the proposed automatic pipeline to filter collected data and construct two subsets named HardMo-Hand and HardMo-Foot. Extensive experiments demonstrate the effectiveness of the annotation pipeline and the data-driven solution to failure cases. Specifically after being trained on HardMo HMR an early pioneering method can even outperform the current state of the art 4DHumans on our benchmarks.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Human Mesh Recovery

Similar Papers 제목 키워드 기반

XmoPipe: A Pipeline for Large-Scale In-the-Wild Human Motion Dataset Construction

2026-06-17 · Nathan Salazar, Emmanuel Dellandréa, Mathieu Lefort, Alexandre Meyer arxiv

Large-scale human motion datasets are essential for training robust motion models for analysis, synthesis, and understanding. While marker-based motion capture provides precise data, it is costly and limited in scale and…

Emotion4MIDI: a Lyrics-based Emotion-Labeled Symbolic Music Dataset

2023-07-27 · Serkan Sulun, Pedro Oliveira, Paula Viana

We present a new large-scale emotion-labeled symbolic music dataset consisting of 12k MIDI songs. To create this dataset, we first trained emotion classification models on the GoEmotions dataset, achieving state-of-the-a…

Emotion Classification

Make-An-Animation: Large-Scale Text-conditional 3D Human Motion Generation

2023-05-16 · ICCV 2023 1 · Samaneh Azadi, Akbar Shah, Thomas Hayes, Devi Parikh 외

Text-guided human motion generation has drawn significant interest because of its impactful applications spanning animation and robotics. Recently, application of diffusion models for motion generation has enabled improv…

Motion GenerationMotion SynthesisText-to-Video GenerationVideo Generation

FUSION: Full-Body Unified Motion Prior for Body and Hands via Diffusion

2026-01-07 · Enes Duran, Nikos Athanasiou, Muhammed Kocabas, Michael J. Black 외 arxiv

Hands are central to interacting with our surroundings and conveying gestures, making their inclusion essential for full-body motion synthesis. Despite this, existing human motion synthesis methods fall short: some ignor…

Motion Synthesis

MotionBank: A Large-scale Video Motion Benchmark with Disentangled Rule-based Annotations

2024-10-17 · Liang Xu, Shaoyang Hua, Zili Lin, Yifan Liu 외

In this paper, we tackle the problem of how to build and benchmark a large motion model (LMM). The ultimate goal of LMM is to serve as a foundation model for versatile motion-related tasks, e.g., human motion generation,…

Caption GenerationMotion Generation