paper-with-me

홈 › Papers

EgoSim: An Egocentric Multi-view Simulator and Real Dataset for Body-worn Cameras during Motion and Activity

2025-02-25 · Dominik Hollidt, Paul Streli, Jiaxi Jiang, Yasaman Haghighi, Changlin Qian, Xintong Liu, Christian Holz

Research on egocentric tasks in computer vision has mostly focused on head-mounted cameras, such as fisheye cameras or embedded cameras inside immersive headsets. We argue that the increasing miniaturization of optical sensors will lead to the prolific integration of cameras into many more body-worn devices at various locations. This will bring fresh perspectives to established tasks in computer vision and benefit key areas such as human motion tracking, body pose estimation, or action recognition -- particularly for the lower body, which is typically occluded. In this paper, we introduce EgoSim, a novel simulator of body-worn cameras that generates realistic egocentric renderings from multiple perspectives across a wearer's body. A key feature of EgoSim is its use of real motion capture data to render motion artifacts, which are especially noticeable with arm- or leg-worn cameras. In addition, we introduce MultiEgoView, a dataset of egocentric footage from six body-worn cameras and ground-truth full-body 3D poses during several activities: 119 hours of data are derived from AMASS motion sequences in four high-fidelity virtual environments, which we augment with 5 hours of real-world motion data from 13 participants using six GoPro cameras and 3D body pose references from an Xsens motion capture suit. We demonstrate EgoSim's effectiveness by training an end-to-end video-only 3D pose estimation network. Analyzing its domain gap, we show that our dataset and simulator substantially aid training for inference on real-world data. EgoSim code & MultiEgoView dataset: https://siplab.org/projects/EgoSim

📄 PDF Abstract BibTeX arXiv:2502.18373

Code (0)

등록된 구현이 없습니다.

Tasks

3D Pose EstimationAction RecognitionPose Estimation

Similar Papers 제목 키워드 기반

EgoSim: Egocentric World Simulator for Embodied Interaction Generation

2026-04-01 · Jinkun Hao, Mingda Jia, Ruiyan Wang, Hongrui Zhu 외 arxiv

We introduce EgoSim, a closed-loop egocentric world simulator that generates spatially consistent interaction videos and persistently updates the underlying 3D scene state for continuous simulation. Existing egocentric s…

Point Clouds

EgoForge: Goal-Directed Egocentric World Simulator

2026-03-20 · Yifan Shen, Jiateng Liu, Xinzhuo Li, Yuanzhe Liu 외 arxiv

Generative world models have shown promise for simulating dynamic environments, yet egocentric video remains challenging due to rapid viewpoint changes, frequent hand-object interactions, and goal-directed procedures who…

PlayerOne: Egocentric World Simulator

2025-06-11 · Yuanpeng Tu, Hao Luo, Xi Chen, Xiang Bai 외

We introduce PlayerOne, the first egocentric realistic world simulator, facilitating immersive and unrestricted exploration within vividly dynamic environments. Given an egocentric scene image from the user, PlayerOne ca…

Video Generation

RetailSMV: Exocentric vs. Egocentric Adaptation of Foundation Video World Models in Retail

2026-07-01 · Amirreza Rouhi, Rajat Aggarwal, Parikshit Sakurikar, Anoop M. Namboodiri 외 arxiv

Foundation video diffusion models are increasingly viewed as world simulators for embodied agents, yet their pretraining on internet-scale generic video leaves them poorly aligned with real-world deployment domains. We s…

EgoInteract: Synthetic Egocentric Videos Generation for Interaction Understanding and Anticipation

2026-05-18 · Rosario Leonardi, Francesco Ragusa, Daniele Materia, Alessandro Passanisi 외 arxiv

Collecting large-scale egocentric video datasets with dense spatial and temporal annotations is costly, slow, and often constrained by environmental biases, privacy constraints, and limited coverage of interaction patter…

Active Object DetectionAction SegmentationVideo Generation