paper-with-me

Papers

Task-Oriented Human-Object Interactions Generation with Implicit Neural Representations

2023-03-23 · Quanzhou Li, Jingbo Wang, Chen Change Loy, Bo Dai

Digital human motion synthesis is a vibrant research field with applications in movies, AR/VR, and video games. Whereas methods were proposed to generate natural and realistic human motions, most only focus on modeling humans and largely ignore object movements. Generating task-oriented human-object interaction motions in simulation is challenging. For different intents of using the objects, humans conduct various motions, which requires the human first to approach the objects and then make them move consistently with the human instead of staying still. Also, to deploy in downstream applications, the synthesized motions are desired to be flexible in length, providing options to personalize the predicted motions for various purposes. To this end, we propose TOHO: Task-Oriented Human-Object Interactions Generation with Implicit Neural Representations, which generates full human-object interaction motions to conduct specific tasks, given only the task type, the object, and a starting human status. TOHO generates human-object motions in three steps: 1) it first estimates the keyframe poses of conducting a task given the task type and object information; 2) then, it infills the keyframes and generates continuous motions; 3) finally, it applies a compact closed-form object motion estimation to generate the object motion. Our method generates continuous motions that are parameterized only by the temporal coordinate, which allows for upsampling or downsampling of the sequence to arbitrary frames and adjusting the motion speeds by designing the temporal coordinate vector. We demonstrate the effectiveness of our method, both qualitatively and quantitatively. This work takes a step further toward general human-scene interaction simulation.

📄 PDF Abstract BibTeX arXiv:2303.13129

Code (0)

등록된 구현이 없습니다.

Tasks

Human-Object Interaction DetectionMotion EstimationMotion SynthesisObject

Similar Papers 제목 키워드 기반

OakInk: A Large-scale Knowledge Repository for Understanding Hand-Object Interaction

2022-03-29 · CVPR 2022 1 · Lixin Yang, Kailin Li, Xinyu Zhan, Fei Wu 외

Learning how humans manipulate objects requires machines to acquire knowledge from two perspectives: one for understanding object affordances and the other for learning human's interactions based on the affordances. Even…

Grasp GenerationObjectPose Estimation

Learning from Observer Gaze:Zero-Shot Attention Prediction Oriented by Human-Object Interaction Recognition

2024-05-16 · Yuchen Zhou, Linkai Liu, Chao Gou

Most existing attention prediction research focuses on salient instances like humans and objects. However, the more complex interaction-oriented attention, arising from the comprehension of interactions between instances…

Human-Object Interaction Detection

Learning from Observer Gaze: Zero-Shot Attention Prediction Oriented by Human-Object Interaction Recognition

2024-01-01 · CVPR 2024 1 · Yuchen Zhou, Linkai Liu, Chao Gou

Most existing attention prediction research focuses on salient instances like humans and objects. However the more complex interaction-oriented attention arising from the comprehension of interactions between instanc…

Human-Object Interaction Detection

AffordPose: A Large-scale Dataset of Hand-Object Interactions with Affordance-driven Hand Pose

2023-09-16 · ICCV 2023 1 · Juntao Jian, Xiuping Liu, Manyi Li, Ruizhen Hu 외

How human interact with objects depends on the functional roles of the target objects, which introduces the problem of affordance-aware hand-object interaction. It requires a large number of human demonstrations for the …

DiversityObject

MMHOI: Modeling Complex 3D Multi-Human Multi-Object Interactions

2025-10-09 · Kaen Kogashi, Anoop Cherian, Meng-Yu Jennifer Kuo arxiv

Real-world scenes often feature multiple humans interacting with multiple objects in ways that are causal, goal-oriented, or cooperative. Yet existing 3D human-object interaction (HOI) benchmarks consider only a fraction…

Action Recognition