paper-with-me

홈 › Papers

NeuralDome: A Neural Modeling Pipeline on Multi-View Human-Object Interactions

2022-12-15 · CVPR 2023 1 · Juze Zhang, Haimin Luo, Hongdi Yang, Xinru Xu, Qianyang Wu, Ye Shi, Jingyi Yu, Lan Xu, Jingya Wang

Humans constantly interact with objects in daily life tasks. Capturing such processes and subsequently conducting visual inferences from a fixed viewpoint suffers from occlusions, shape and texture ambiguities, motions, etc. To mitigate the problem, it is essential to build a training dataset that captures free-viewpoint interactions. We construct a dense multi-view dome to acquire a complex human object interaction dataset, named HODome, that consists of $\sim$75M frames on 10 subjects interacting with 23 objects. To process the HODome dataset, we develop NeuralDome, a layer-wise neural processing pipeline tailored for multi-view video inputs to conduct accurate tracking, geometry reconstruction and free-view rendering, for both human subjects and objects. Extensive experiments on the HODome dataset demonstrate the effectiveness of NeuralDome on a variety of inference, modeling, and rendering tasks. Both the dataset and the NeuralDome tools will be disseminated to the community for further development.

📄 PDF Abstract BibTeX arXiv:2212.07626

Code (0)

등록된 구현이 없습니다.

Tasks

Human-Object Interaction Detection

Similar Papers 제목 키워드 기반

DCHM: Depth-Consistent Human Modeling for Multiview Detection

2025-07-19 · Jiahao Ma, Tianyu Wang, Miaomiao Liu, David Ahmedt-Aristizabal 외 arxiv

Multiview pedestrian detection typically involves two stages: human modeling and pedestrian localization. Human modeling represents pedestrians in 3D space by fusing multiview information, making its quality crucial for …

Pedestrian DetectionMultiview DetectionDepth EstimationPoint Clouds

ReImagine: Rethinking Controllable High-Quality Human Video Generation via Image-First Synthesis

2026-04-21 · Zhengwentai Sun, Keru Zheng, Chenghong Li, Hongjie Liao 외 arxiv

Human video generation remains challenging due to the difficulty of jointly modeling human appearance, motion, and camera viewpoint under limited multi-view data. Existing methods often address these factors separately, …

Video GenerationImage Generation

HumanDreamer-X: Photorealistic Single-image Human Avatars Reconstruction via Gaussian Restoration

2025-04-04 · Boyuan Wang, Runqi Ouyang, XiaoFeng Wang, Zheng Zhu 외

Single-image human reconstruction is vital for digital human modeling applications but remains an extremely challenging task. Current approaches rely on generative models to synthesize multi-view images for subsequent 3D…

3DGS3D Reconstruction

Smart Director: An Event-Driven Directing System for Live Broadcasting

2022-01-11 · Yingwei Pan, Yue Chen, Qian Bao, Ning Zhang 외

Live video broadcasting normally requires a multitude of skills and expertise with domain knowledge to enable multi-camera productions. As the number of cameras keep increasing, directing a live sports broadcast has now …

Event DetectionHighlight Detection

GroomLight: Hybrid Inverse Rendering for Relightable Human Hair Appearance Modeling

2025-03-13 · CVPR 2025 1 · Yang Zheng, Menglei Chai, Delio Vicini, Yuxiao Zhou 외

We present GroomLight, a novel method for relightable hair appearance modeling from multi-view images. Existing hair capture methods struggle to balance photorealistic rendering with relighting capabilities. Analytical m…

Inverse RenderingNeural Rendering