paper-with-me

Papers

Closely Interactive Human Reconstruction with Proxemics and Physics-Guided Adaption

2024-04-17 · CVPR 2024 1 · Buzhen Huang, Chen Li, Chongyang Xu, Liang Pan, Yangang Wang, Gim Hee Lee

Existing multi-person human reconstruction approaches mainly focus on recovering accurate poses or avoiding penetration, but overlook the modeling of close interactions. In this work, we tackle the task of reconstructing closely interactive humans from a monocular video. The main challenge of this task comes from insufficient visual information caused by depth ambiguity and severe inter-person occlusion. In view of this, we propose to leverage knowledge from proxemic behavior and physics to compensate the lack of visual information. This is based on the observation that human interaction has specific patterns following the social proxemics. Specifically, we first design a latent representation based on Vector Quantised-Variational AutoEncoder (VQ-VAE) to model human interaction. A proxemics and physics guided diffusion model is then introduced to denoise the initial distribution. We design the diffusion model as dual branch with each branch representing one individual such that the interaction can be modeled via cross attention. With the learned priors of VQ-VAE and physical constraint as the additional information, our proposed approach is capable of estimating accurate poses that are also proxemics and physics plausible. Experimental results on Hi4D, 3DPW, and CHI3D demonstrate that our method outperforms existing approaches. The code is available at \url{https://github.com/boycehbz/HumanInteraction}.

📄 PDF Abstract BibTeX arXiv:2404.11291

Code (1)

boycehbz/humaninteraction 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

VQ-VAE VQ-VAE is a type of variational autoencoder that uses vector quantisation to obtain a discrete latent representation. It differs from…
Focus 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Generative Proxemics: A Prior for 3D Social Interaction from Images

2023-06-15 · CVPR 2024 1 · Lea Müller, Vickie Ye, Georgios Pavlakos, Michael Black 외

Social interaction is a fundamental aspect of human behavior and communication. The way individuals position themselves in relation to others, also known as proxemics, conveys social cues and affects the dynamics of soci…

3D ReconstructionDenoisingSingle-View 3D Reconstruction

Reconstructing Close Human Interaction with Appearance and Proxemics Reasoning

2025-07-03 · CVPR 2025 1 · Buzhen Huang, Chen Li, Chongyang Xu, Dongyue Lu 외

Due to visual ambiguities and inter-person occlusions, existing human pose estimation methods cannot recover plausible close interactions from in-the-wild videos. Even state-of-the-art large foundation models~(\eg, SAM) …

Pose Estimation

Towards Socially Compliant Navigation in Deep Reinforcement Learning via Proxemics-Based Reward Modeling

2026-08-13 · Takieddine Soualhi, Jacques Saraydaryan, Laetitia Matignon arxiv

Developing effective robot navigation methods in crowded environments is essential for real-world applications. Although recent deep reinforcement learning (DRL) methods have improved navigation performance in crowded en…

Reinforcement LearningRobot Navigation

SimuScene: Simulation-Ready Compositional 3D Scene Reconstruction from a Single Image

2026-06-02 · Inhee Lee, Sangwon Baik, Sungjoo Kim, Hyeonwoo Kim 외 arxiv

Reconstructing interactive, simulation-ready 3D scenes from a single image is a critical bottleneck for robotic manipulation. While recent single-image lifters recover plausible per-object shapes, composing them yields s…

3D Reconstruction

Tracking Motion and Proxemics using Thermal-sensor Array

2015-11-25 · Chandrayee Basu, Anthony Rowe

Indoor tracking has all-pervasive applications beyond mere surveillance, for example in education, health monitoring, marketing, energy management and so on. Image and video based tracking systems are intrusive. Thermal …

energy managementManagementMarketingMotion Detection