paper-with-me

Papers

Gaze-guided Hand-Object Interaction Synthesis: Dataset and Method

2024-03-24 · Jie Tian, Ran Ji, Lingxiao Yang, Suting Ni, Yuexin Ma, Lan Xu, Jingyi Yu, Ye Shi, Jingya Wang

Gaze plays a crucial role in revealing human attention and intention, particularly in hand-object interaction scenarios, where it guides and synchronizes complex tasks that require precise coordination between the brain, hand, and object. Motivated by this, we introduce a novel task: Gaze-Guided Hand-Object Interaction Synthesis, with potential applications in augmented reality, virtual reality, and assistive technologies. To support this task, we present GazeHOI, the first dataset to capture simultaneous 3D modeling of gaze, hand, and object interactions. This task poses significant challenges due to the inherent sparsity and noise in gaze data, as well as the need for high consistency and physical plausibility in generating hand and object motions. To tackle these issues, we propose a stacked gaze-guided hand-object interaction diffusion model, named GHO-Diffusion. The stacked design effectively reduces the complexity of motion generation. We also introduce HOI-Manifold Guidance during the sampling stage of GHO-Diffusion, enabling fine-grained control over generated motions while maintaining the data manifold. Additionally, we propose a spatial-temporal gaze feature encoding for the diffusion condition and select diffusion results based on consistency scores between gaze-contact maps and gaze-interaction trajectories. Extensive experiments highlight the effectiveness of our method and the unique contributions of our dataset. More details in https://takiee.github.io/gaze-hoi/.

📄 PDF Abstract BibTeX arXiv:2403.16169

Code (0)

등록된 구현이 없습니다.

Tasks

DenoisingHuman motion predictionMotion Generationmotion predictionObject

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

G3Ego: Gaze-Guided Graphs for Egocentric Action Understanding

2026-08-20 · Marko Haralović, Akash Ramakrishnan, Estefania Talavera Martinez arxiv

Egocentric action understanding is often addressed using large video models pretrained on extensive exocentric datasets. However, many first-person actions depend on a small number of hand-object interactions involving o…

Action UnderstandingAction Recognition

Vision-Guided Action: Enhancing 3D Human Motion Prediction with Gaze-informed Affordance in 3D Scenes

2025-01-01 · CVPR 2025 1 · Ting Yu, Yi Lin, Jun Yu, Zhenyu Lou 외

Recent advances in human motion prediction (HMP) have shifted focus from isolated motion data to integrating human-scene correlations. In particular, the latest methods leverage human gaze points, using their spatial…

Human motion predictionHuman-Object Interaction Detectionmotion prediction

Object Motion Guided Human Motion Synthesis

2023-09-28 · Jiaman Li, Jiajun Wu, C. Karen Liu

Modeling human behaviors in contextual environments has a wide range of applications in character animation, embodied AI, VR/AR, and robotics. In real-world scenarios, humans frequently interact with the environment and …

DenoisingHuman-Object Interaction DetectionMotion SynthesisObject

A Hands-free Spatial Selection and Interaction Technique using Gaze and Blink Input with Blink Prediction for Extended Reality

2025-01-20 · Tim Rolff, Jenny Gabel, Lauren Zerbin, Niklas Hypki 외

Gaze-based interaction techniques have created significant interest in the field of spatial interaction. Many of these methods require additional input modalities, such as hand gestures (e.g., gaze coupled with pinch). T…

ChildPlay-Hand: A Dataset of Hand Manipulations in the Wild

2024-09-14 · Arya Farkhondeh, Samy Tafasca, Jean-Marc Odobez

Hand-Object Interaction (HOI) is gaining significant attention, particularly with the creation of numerous egocentric datasets driven by AR/VR applications. However, third-person view HOI has received less attention, esp…

Action RecognitionHand DetectionObject