paper-with-me

홈 › Papers

Policy-as-Data: Learning Generalizable HOI Diffusion Models from Simulated Physics

2026-06-22 · Shujia Li, Jianshu Hu, Haiyu Zhang, Yunpeng Jiang, Haoyuan Jin, Xinyuan Chen, Yaohui Wang, Yutong Ban arxiv

Synthesizing realistic Human-Object Interactions (HOI) is critical for creating embodied avatars and functional virtual environments. However, current data-driven approaches primarily rely on motion capture datasets, which are expensive to scale and limited in functional diversity. Models trained with these datasets fail to generalize to unseen objects and maintain physical consistency over long horizons. In this paper, we propose a novel framework that leverages a physics simulator to overcome the data-scarcity bottleneck in HOI generation. Specifically, we propose a scalable pipeline, called \ours, which leverages policies trained with reinforcement learning in a physics simulator for task-oriented data generation and trains a generative model on the augmented dataset for generalizable HOI generation. To seamlessly utilize the synthetic data, we introduce a coarse-to-fine retargeting process that bridges the representation gap between the simplified model used in physics simulator and the standard parametric body models required for generative training. Validated through comprehensive experiments, our method demonstrates enhanced generalization to unseen objects and the capability of long-horizon generation, while exhibiting greater dynamic diversity and physical plausibility.

📄 PDF Abstract BibTeX arXiv:2606.22806

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Planning-Guided Diffusion Policy Learning for Generalizable Contact-Rich Bimanual Manipulation

2024-12-03 · Xuanlin Li, Tong Zhao, Xinghao Zhu, Jiuguang Wang 외

Contact-rich bimanual manipulation involves precise coordination of two arms to change object states through strategically selected contacts and motions. Due to the inherent complexity of these tasks, acquiring sufficien…

Data Augmentation

GenORM: Generalizable One-shot Rope Manipulation with Parameter-Aware Policy

2023-06-14 · So Kuroki, Jiaxian Guo, Tatsuya Matsushima, Takuya Okubo 외

Due to the inherent uncertainty in their deformability during motion, previous methods in rope manipulation often require hundreds of real-world demonstrations to train a manipulation policy for each rope, even for simpl…

Iterative Distillation for Reward-Guided Fine-Tuning of Diffusion Models in Biomolecular Design

2025-07-01 · Xingyu Su, Xiner Li, Masatoshi Uehara, Sunwoo Kim 외

We address the problem of fine-tuning diffusion models for reward-guided generation in biomolecular design. While diffusion models have proven highly effective in modeling complex, high-dimensional data distributions, re…

EBT-Policy: Energy Unlocks Emergent Physical Reasoning Capabilities

2025-10-31 · Travis Davies, Yiqi Huang, Alexi Gladstone, Yunxin Liu 외 arxiv

Implicit policies parameterized by generative models, such as Diffusion Policy, have become the standard for policy learning and Vision-Language-Action (VLA) models in robotics. However, these approaches often suffer fro…

GenDOM: Generalizable One-shot Deformable Object Manipulation with Parameter-Aware Policy

2023-09-16 · So Kuroki, Jiaxian Guo, Tatsuya Matsushima, Takuya Okubo 외

Due to the inherent uncertainty in their deformability during motion, previous methods in deformable object manipulation, such as rope and cloth, often required hundreds of real-world demonstrations to train a manipulati…

Deformable Object ManipulationObject