paper-with-me

Papers

HandDiffuse: Generative Controllers for Two-Hand Interactions via Diffusion Models

2023-12-08 · Pei Lin, Sihang Xu, Hongdi Yang, Yiran Liu, Xin Chen, Jingya Wang, Jingyi Yu, Lan Xu

Existing hands datasets are largely short-range and the interaction is weak due to the self-occlusion and self-similarity of hands, which can not yet fit the need for interacting hands motion generation. To rescue the data scarcity, we propose HandDiffuse12.5M, a novel dataset that consists of temporal sequences with strong two-hand interactions. HandDiffuse12.5M has the largest scale and richest interactions among the existing two-hand datasets. We further present a strong baseline method HandDiffuse for the controllable motion generation of interacting hands using various controllers. Specifically, we apply the diffusion model as the backbone and design two motion representations for different controllers. To reduce artifacts, we also propose Interaction Loss which explicitly quantifies the dynamic interaction process. Our HandDiffuse enables various applications with vivid two-hand interactions, i.e., motion in-betweening and trajectory control. Experiments show that our method outperforms the state-of-the-art techniques in motion generation and can also contribute to data augmentation for other datasets. Our dataset, corresponding codes, and pre-trained models will be disseminated to the community for future research towards two-hand interaction modeling.

📄 PDF Abstract BibTeX arXiv:2312.04867

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationMotion Generationmotion in-betweeningTemporal Sequences

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

When Generative AI Meets Extended Reality: Enabling Scalable and Natural Interactions

2026-01-13 · Mingyu Zhu, Jiangong Chen, Bin Li arxiv

Extended Reality (XR), including virtual, augmented, and mixed reality, provides immersive and interactive experiences across diverse applications, from VR-based education to AR-based assistance and MR-based training. Ho…

GraspDiffusion: Synthesizing Realistic Whole-body Hand-Object Interaction

2024-10-17 · Patrick Kwon, Hanbyul Joo

Recent generative models can synthesize high-quality images but often fail to generate humans interacting with objects using their hands. This arises mostly from the model's misunderstanding of such interactions, and the…

Human-Object Interaction DetectionImage GenerationObject

DC-ControlNet: Decoupling Inter- and Intra-Element Conditions in Image Generation with Diffusion Models

2025-02-20 · Hongji Yang, Wencheng Han, Yucheng Zhou, Jianbing Shen

In this paper, we introduce DC (Decouple)-ControlNet, a highly flexible and precisely controllable framework for multi-condition image generation. The core idea behind DC-ControlNet is to decouple control conditions, tra…

Conditional Image GenerationImage Generation

Affordance-Guided Diffusion Prior for 3D Hand Reconstruction

2025-10-01 · Naru Suzuki, Takehiko Ohkawa, Tatsuro Banno, Jihyun Lee 외 arxiv

How can we reconstruct 3D hand poses when large portions of the hand are heavily occluded by itself or by objects? Humans often resolve such ambiguities by leveraging contextual knowledge -- such as affordances, where an…

Hand Pose Estimation

Affordance Diffusion: Synthesizing Hand-Object Interactions

2023-03-21 · CVPR 2023 1 · Yufei Ye, Xueting Li, Abhinav Gupta, Shalini De Mello 외

Recent successes in image synthesis are powered by large-scale diffusion models. However, most methods are currently limited to either text- or image-conditioned generation for synthesizing an entire image, texture trans…

DescriptiveImage GenerationObject