paper-with-me

Papers

UniCon3R: Unified Contact-aware 4D Human-Scene Reconstruction from Monocular Video

2026-04-21 · Tanuj Sur, Shashank Tripathi, Nikos Athanasiou, Ha Linh Nguyen, Kai Xu, Michael J. Black, Angela Yao arxiv

We introduce UniCon3R, a unified feed-forward framework for online human-scene 4D reconstruction from monocular video. Current feed-forward human-scene reconstruction methods suffer from artifacts, where bodies float above the ground or penetrate parts of the scene. A key reason is the lack of effective interaction modelling between the human and the environment. Our goal is to exploit contact between the human and the scene during inference to actively improve the human mesh reconstruction. To that end, we explicitly model interaction by inferring 4D contact from the human pose and scene geometry and use the contact as a corrective cue for generating the pose. This enables UniCon3R to jointly recover scene geometry and spatially aligned 4D humans within the scene. Experiments on standard human-centric video benchmarks show that UniCon3R outperforms state-of-the-art baselines on physical plausibility and global human motion estimation while preserving fast, feed-forward inference speeds. The results validate our central claim: contact serves as a powerful internal prior, thus establishing a new paradigm for physically grounded joint human-scene reconstruction. Project page is available at https://surtantheta.github.io/UniCon3R .

📄 PDF Abstract BibTeX arXiv:2604.19923

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

UniCon+: ICTCAS-UCAS Submission to the AVA-ActiveSpeaker Task at ActivityNet Challenge 2022

2022-06-22 · Yuanhang Zhang, Susan Liang, Shuang Yang, Shiguang Shan

This report presents a brief description of our winning solution to the AVA Active Speaker Detection (ASD) task at ActivityNet Challenge 2022. Our underlying model UniCon+ continues to build on our previous work, the Uni…

Active Speaker DetectionAudio-Visual Active Speaker Detection

UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild

2023-05-18 · NeurIPS 2023 11 · Can Qin, Shu Zhang, Ning Yu, Yihao Feng 외

Achieving machine autonomy and human control often represent divergent objectives in the design of interactive AI systems. Visual generative foundation models such as Stable Diffusion show promise in navigating these goa…

Image Generation

UniConFlow: A Unified Constrained Generalization Framework for Certified Motion Planning with Flow Matching Models

2025-06-03 · Zewen Yang, Xiaobing Dai, Dian Yu, Qianru Li 외

Generative models have become increasingly powerful tools for robot motion generation, enabling flexible and multimodal trajectory generation across various tasks. Yet, most existing approaches remain limited in handling…

Collision AvoidanceMotion GenerationMotion Planning

UniCon: Unified Framework for Efficient Contrastive Alignment via Kernels

2026-04-17 · Hangke Sui, Yuqing Wang, Minh N Do arxiv

Contrastive objectives power state-of-the-art multimodal models, but their training remains slow, relying on long stochastic optimization. We propose a Unified Framework for Efficient Contrastive Alignment via Kernels (U…

Stochastic Optimization

HUGS: Guiding Unified Dexterous Grasp Synthesis Across Modes and Scales via Learned Human Priors

2026-07-06 · Mingrui Yu, Yongpeng Jiang, Yongyi Jia, Kangchen Lv 외 arxiv

Dexterous grasping across diverse object scales requires contact modes ranging from two-finger pinches to bimanual grasps. Existing dexterous grasp synthesis methods reduce the high-dimensional optimization space with ma…