paper-with-me

Papers

Surface Keypoint Representation for Multi-Object and Articulated Human-Object Interaction Generation

2026-08-04 · Xiaogang Peng, Zeyu Han, Zichong Meng, Yiming Xie, Jihua Zhu, Gang Hua, Huaizu Jiang arxiv

Daily activities require humans to coordinate whole-body motion with the motion of surrounding objects. Despite recent progress in human-object interaction (HOI) generation, most existing methods assume interactions with a single rigid object and do not extend well to scenarios involving a variable number of objects or articulated objects with diverse joint mechanisms. We propose surface keypoint trajectories as an object motion representation: for each rigid component, whether a standalone object or one part of an articulated assembly, we track a small set of non-collinear surface points over time. This representation handles multi-object coordination and diverse articulation mechanisms directly from point dynamics without requiring explicit joint-type specification. To model when and where each body region contacts each object, we introduce a spatio-temporal contact distance field that extends distance-based contact modeling to whole-body, multi-object, and articulated settings. We factorize HOI generation into three stages: generating object motions from text or waypoints, predicting the contact distance field, and synthesizing whole-body motion with contact-guided optimization. Experiments on ParaHome, HIMO, ARCTIC, and OMOMO demonstrate better or comparable performance to existing methods across single-object, multi-object, and articulated interaction settings.

📄 PDF Abstract BibTeX arXiv:2608.03158

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Skeleton-Driven Neural Occupancy Representation for Articulated Hands

2021-09-23 · Korrawe Karunratanakul, Adrian Spurr, Zicong Fan, Otmar Hilliges 외

We present Hand ArticuLated Occupancy (HALO), a novel representation of articulated hands that bridges the advantages of 3D keypoints and neural implicit surfaces and can be used in end-to-end trainable architectures. Un…

ViSER: Video-Specific Surface Embeddings for Articulated 3D Shape Reconstruction

2021-12-01 · NeurIPS 2021 12 · Gengshan Yang, Deqing Sun, Varun Jampani, Daniel Vlasic 외

We introduce ViSER, a method for recovering articulated 3D shapes and dense3D trajectories from monocular videos. Previous work on high-quality reconstruction of dynamic 3D shapes typically relies on multiple camera vie…

3D Shape Reconstruction from Videos

REArtGS: Reconstructing and Generating Articulated Objects via 3D Gaussian Splatting with Geometric and Motion Constraints

2025-03-09 · Di wu, Liu Liu, Zhou Linli, Anran Huang 외

Articulated objects, as prevalent entities in human life, their 3D representations play crucial roles across various applications. However, achieving both high-fidelity textured surface reconstruction and dynamic generat…

Surface Reconstruction

Learning Articulated Shape with Keypoint Pseudo-labels from Web Images

2023-04-27 · CVPR 2023 1 · Anastasis Stathopoulos, Georgios Pavlakos, Ligong Han, Dimitris Metaxas

This paper shows that it is possible to learn models for monocular 3D reconstruction of articulated objects (e.g., horses, cows, sheep), using as few as 50-150 images labeled with 2D keypoints. Our proposed approach invo…

3D ReconstructionKeypoint EstimationObject

Learning Implicit Representation for Reconstructing Articulated Objects

2024-01-16 · Hao Zhang, Fang Li, Samyak Rawlekar, Narendra Ahuja

3D Reconstruction of moving articulated objects without additional information about object structure is a challenging problem. Current methods overcome such challenges by employing category-specific skeletal models. Con…

3D ReconstructionObject