paper-with-me

홈 › Papers

Moving by Looking: Towards Vision-Driven Avatar Motion Generation

2025-09-23 · Markos Diomataris, Berat Mert Albaba, Giorgio Becherini, Partha Ghosh, Omid Taheri, Michael J. Black arxiv

The way we perceive the world fundamentally shapes how we move, whether it is how we navigate in a room or how we interact with other humans. Current human motion generation methods, neglect this interdependency and use task-specific ``perception'' that differs radically from that of humans. We argue that the generation of human-like avatar behavior requires human-like perception. Consequently, in this work we present CLOPS, the first human avatar that solely uses egocentric vision to perceive its surroundings and navigate. Using vision as the primary driver of motion however, gives rise to a significant challenge for training avatars: existing datasets have either isolated human motion, without the context of a scene, or lack scale. We overcome this challenge by decoupling the learning of low-level motion skills from learning of high-level control that maps visual input to motion. First, we train a motion prior model on a large motion capture dataset. Then, a policy is trained using Q-learning to map egocentric visual inputs to high-level control commands for the motion prior. Our experiments empirically demonstrate that egocentric vision can give rise to human-like motion characteristics in our avatars. For example, the avatars walk such that they avoid obstacles present in their visual field. These findings suggest that equipping avatars with human-like sensors, particularly egocentric vision, holds promise for training avatars that behave like humans.

📄 PDF Abstract BibTeX arXiv:2509.19259

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AvatarCLIP: Zero-Shot Text-Driven Generation and Animation of 3D Avatars

2022-05-17 · Fangzhou Hong, Mingyuan Zhang, Liang Pan, Zhongang Cai 외

3D avatar creation plays a crucial role in the digital age. However, the whole production process is prohibitively time-consuming and labor-intensive. To democratize this technology to a larger audience, we propose Avata…

3D geometryLanguage ModellingMotion SynthesisTexture Synthesis

Drivable Avatar Clothing: Faithful Full-Body Telepresence with Dynamic Clothing Driven by Sparse RGB-D Input

2023-10-09 · Donglai Xiang, Fabian Prada, Zhe Cao, Kaiwen Guo 외

Clothing is an important part of human appearance but challenging to model in photorealistic avatars. In this work we present avatars with dynamically moving loose clothing that can be faithfully driven by sparse RGB-D i…

Moving Avatars and Agents in Social Extended Reality Environments

2023-06-26 · Jann Philipp Freiwald, Susanne Schmidt, Bernhard E. Riecke, Frank Steinicke

Natural interaction between multiple users within a shared virtual environment (VE) relies on each other's awareness of the current position of the interaction partners. This, however, cannot be warranted when users empl…

Navigate

Splat-Portrait: Generalizing Talking Heads with Gaussian Splatting

2026-01-26 · Tong Shi, Melonie de Almeida, Daniela Ivanova, Nicolas Pugeault 외 arxiv

Talking Head Generation aims at synthesizing natural-looking talking videos from speech and a single portrait image. Previous 3D talking head generation methods have relied on domain-specific heuristics such as warping-b…

Talking Head GenerationNovel View Synthesis3D ReconstructionMotion Synthesis

Deblur-Avatar: Animatable Avatars from Motion-Blurred Monocular Videos

2025-01-23 · Xianrui Luo, Juewen Peng, Zhongang Cai, Lei Yang 외

We introduce Deblur-Avatar, a novel framework for modeling high-fidelity, animatable 3D human avatars from motion-blurred monocular video inputs. Motion blur is prevalent in real-world dynamic video capture, especially d…

3DGS