paper-with-me

Papers

Generative Head-Mounted Camera Captures for Photorealistic Avatars

2025-07-08 · Shaojie Bai, Seunghyeon Seo, Yida Wang, Chenghui Li, Owen Wang, Te-Li Wang, Tianyang Ma, Jason Saragih, Shih-En Wei, Nojun Kwak, Hyung Jun Kim

Enabling photorealistic avatar animations in virtual and augmented reality (VR/AR) has been challenging because of the difficulty of obtaining ground truth state of faces. It is physically impossible to obtain synchronized images from head-mounted cameras (HMC) sensing input, which has partial observations in infrared (IR), and an array of outside-in dome cameras, which have full observations that match avatars' appearance. Prior works relying on analysis-by-synthesis methods could generate accurate ground truth, but suffer from imperfect disentanglement between expression and style in their personalized training. The reliance of extensive paired captures (HMC and dome) for the same subject makes it operationally expensive to collect large-scale datasets, which cannot be reused for different HMC viewpoints and lighting. In this work, we propose a novel generative approach, Generative HMC (GenHMC), that leverages large unpaired HMC captures, which are much easier to collect, to directly generate high-quality synthetic HMC images given any conditioning avatar state from dome captures. We show that our method is able to properly disentangle the input conditioning signal that specifies facial expression and viewpoint, from facial appearance, leading to more accurate ground truth. Furthermore, our method can generalize to unseen identities, removing the reliance on the paired captures. We demonstrate these breakthroughs by both evaluating synthetic HMC images and universal face encoders trained from these new HMC-avatar correspondences, which achieve better data efficiency and state-of-the-art accuracy.

📄 PDF Abstract BibTeX arXiv:2507.05620

Code (0)

등록된 구현이 없습니다.

Tasks

Disentanglement

Similar Papers 제목 키워드 기반

Real-Time Simulated Avatar from Head-Mounted Sensors

2024-03-11 · CVPR 2024 1 · Zhengyi Luo, Jinkun Cao, Rawal Khirodkar, Alexander Winkler 외

We present SimXR, a method for controlling a simulated avatar from information (headset pose and cameras) obtained from AR / VR headsets. Due to the challenging viewpoint of head-mounted cameras, the human body is often …

Egocentric Pose EstimationHumanoid ControlPose Estimation

Audio Driven Real-Time Facial Animation for Social Telepresence

2025-10-01 · Jiye Lee, Chenghui Li, Linh Tran, Shih-En Wei 외 arxiv

We present an audio-driven real-time system for animating photorealistic 3D facial avatars with minimal latency, designed for social interactions in virtual reality for anyone. Central to our approach is an encoder model…

Revisiting an Old Perspective Projection for Monocular 3D Morphable Models Regression

2026-03-05 · Toby Chong, Ryota Nakajima arxiv

We introduce a novel camera model for monocular 3D Morphable Model (3DMM) regression methods that effectively captures the perspective distortion effect commonly seen in close-up facial images. Fitting 3D morphable model…

EgoRenderer: Rendering Human Avatars from Egocentric Camera Images

2021-11-24 · ICCV 2021 10 · Tao Hu, Kripasindhu Sarkar, Lingjie Liu, Matthias Zwicker 외

We present EgoRenderer, a system for rendering full-body neural avatars of a person captured by a wearable, egocentric fisheye camera that is mounted on a cap or a VR headset. Our system renders photorealistic novel view…

Texture SynthesisTranslation

Recognizing Activities of Daily Living with a Wrist-mounted Camera

2015-11-20 · CVPR 2016 6 · Katsunori Ohnishi, Atsushi Kanehira, Asako Kanezaki, Tatsuya Harada

We present a novel dataset and a novel algorithm for recognizing activities of daily living (ADL) from a first-person wearable camera. Handled objects are crucially important for egocentric ADL recognition. For specific …

object-detectionObject Detection