paper-with-me

Papers

Shape My Moves: Text-Driven Shape-Aware Synthesis of Human Motions

2025-04-04 · CVPR 2025 1 · Ting-Hsuan Liao, Yi Zhou, Yu Shen, Chun-Hao Paul Huang, Saayan Mitra, Jia-Bin Huang, Uttaran Bhattacharya

We explore how body shapes influence human motion synthesis, an aspect often overlooked in existing text-to-motion generation methods due to the ease of learning a homogenized, canonical body shape. However, this homogenization can distort the natural correlations between different body shapes and their motion dynamics. Our method addresses this gap by generating body-shape-aware human motions from natural language prompts. We utilize a finite scalar quantization-based variational autoencoder (FSQ-VAE) to quantize motion into discrete tokens and then leverage continuous body shape information to de-quantize these tokens back into continuous, detailed motion. Additionally, we harness the capabilities of a pretrained language model to predict both continuous shape parameters and motion tokens, facilitating the synthesis of text-aligned motions and decoding them into shape-aware motions. We evaluate our method quantitatively and qualitatively, and also conduct a comprehensive perceptual study to demonstrate its efficacy in generating shape-aware motions.

📄 PDF Abstract BibTeX arXiv:2504.03639

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingMotion GenerationMotion SynthesisQuantization

Similar Papers 제목 키워드 기반

Shape-aware Text-driven Layered Video Editing

2023-01-30 · CVPR 2023 1 · Yao-Chih Lee, Ji-Ze Genevieve Jang, Yi-Ting Chen, Elizabeth Qiu 외

Temporal consistency is essential for video editing applications. Existing work on layered representation of videos allows propagating edits consistently to each frame. These methods, however, can only edit object appear…

Video Editing

SOPHY: Learning to Generate Simulation-Ready Objects with Physical Materials

2025-04-17 · Junyi Cao, Evangelos Kalogerakis

We present SOPHY, a generative model for 3D physics-aware shape synthesis. Unlike existing 3D generative models that focus solely on static geometry or 4D models that produce physics-agnostic animations, our method joint…

Image Reconstruction

ShapeWords: Guiding Text-to-Image Synthesis with 3D Shape-Aware Prompts

2024-12-03 · CVPR 2025 1 · Dmitry Petrov, Pradyumn Goyal, Divyansh Shivashok, Yuanming Tao 외

We introduce ShapeWords, an approach for synthesizing images based on 3D shape guidance and text prompts. ShapeWords incorporates target 3D shape information within specialized tokens embedded together with the input tex…

Image Generation

Shape-Aware Oriented Bounding Box (OBB) to Horizontal Bounding Box (HBB) Conversion

2026-08-06 · Badha Rathna Sabhapathy, Gotam Dahiya, Vishesh Vatsal arxiv

Accurate object detection in aerial and satellite imagery is dependent upon the bounding box representation. This is especially true for spatially oriented objects such as ships or aircrafts. Oriented Bounding Boxes (OBB…

Object Detection

TapMo: Shape-aware Motion Generation of Skeleton-free Characters

2023-10-19 · Jiaxu Zhang, Shaoli Huang, Zhigang Tu, Xin Chen 외

Previous motion generation methods are limited to the pre-rigged 3D human model, hindering their applications in the animation of various non-rigged characters. In this work, we present TapMo, a Text-driven Animation Pip…

Motion Generation