paper-with-me

Papers

Unleashing Guidance Without Classifiers for Human-Object Interaction Animation

2026-03-26 · Ziyin Wang, Sirui Xu, Chuan Guo, Bing Zhou, Jiangshan Gong, Jian Wang, Yu-Xiong Wang, Liang-Yan Gui arxiv

Generating realistic human-object interaction (HOI) animations remains challenging because it requires jointly modeling dynamic human actions and diverse object geometries. Prior diffusion-based approaches often rely on hand-crafted contact priors or human-imposed kinematic constraints to improve contact quality. We propose LIGHT, a data-driven alternative in which guidance emerges from the denoising pace itself, reducing dependence on manually designed priors. Building on diffusion forcing, we factor the representation into modality-specific components and assign individualized noise levels with asynchronous denoising schedules. In this paradigm, cleaner components guide noisier ones through cross-attention, yielding guidance without auxiliary classifiers. We find that this data-driven guidance is inherently contact-aware, and can be enhanced when training is augmented with a broad spectrum of synthetic object geometries, encouraging invariance of contact semantics to geometric diversity. Extensive experiments show that pace-induced guidance more effectively mirrors the benefits of contact priors than conventional classifier-free guidance, while achieving higher contact fidelity, more realistic HOI generation, and stronger generalization to unseen objects and tasks.

📄 PDF Abstract BibTeX arXiv:2603.25734

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Diffusion Classifier Guidance for Non-robust Classifiers

2025-07-01 · Philipp Vaeth, Dibyanshu Kumar, Benjamin Paassen, Magda Gregorová arxiv

Classifier guidance is intended to steer a diffusion process such that a given classifier reliably recognizes the generated data point as a certain class. However, most classifier guidance approaches are restricted to ro…

Stochastic Optimization

CG-HOI: Contact-Guided 3D Human-Object Interaction Generation

2023-11-27 · CVPR 2024 1 · Christian Diller, Angela Dai

We propose CG-HOI, the first method to address the task of generating dynamic 3D human-object interactions (HOIs) from text. We model the motion of both human and object in an interdependent fashion, as semantically rich…

Human-Object Interaction DetectionHuman-Object Interaction GenerationObject

Teaching with Uncertainty: Unleashing the Potential of Knowledge Distillation in Object Detection

2024-06-11 · Junfei Yi, Jianxu Mao, Tengfei Liu, Mingjie Li 외

Knowledge distillation (KD) is a widely adopted and effective method for compressing models in object detection tasks. Particularly, feature-based distillation methods have shown remarkable performance. Existing approach…

Knowledge Distillationobject-detectionObject DetectionTransfer Learning

Diorama: Unleashing Zero-shot Single-view 3D Scene Modeling

2024-11-29 · Qirui Wu, Denys Iliash, Daniel Ritchie, Manolis Savva 외

Reconstructing structured 3D scenes from RGB images using CAD objects unlocks efficient and compact scene representations that maintain compositionality and interactability. Existing works propose training-heavy methods …

3D Shape RetrievalPose Estimation

A Unit Enhancement and Guidance Framework for Audio-Driven Avatar Video Generation

2025-05-06 · S. Z. Zhou, Y. B. Wang, J. F. Wu, T. Hu 외

Audio-driven human animation technology is widely used in human-computer interaction, and the emergence of diffusion models has further advanced its development. Currently, most methods rely on multi-stage generation and…

Human AnimationVideo Generation