paper-with-me

홈 › Papers

Split and Drive: Dual-Axis Disentanglement for Real-Time Gaussian Head Avatars

2026-07-30 · MD Wahiduzzaman Khan, Mingshan Jia, Xiaolin Zhang, En Yu, Kaska Musial-Gabrys arxiv

Creating photorealistic animatable head avatars from a single image remains a fundamental challenge in digital human synthesis. While recent 3D Gaussian Splatting methods have achieved promising results, they rely on external tracking pipelines whose latency is excluded from inference measurements. Furthermore, they adopt unified representations that entangle geometrically distinct facial regions, limiting both expressiveness and rendering fidelity. We propose SpiD (Split and Drive), a single-image Gaussian head avatar framework built on two disentanglement axes. The compute axis internalizes per-frame driving, eliminating external tracking dependency at inference. The feature axis decomposes the avatar into three specialized Gaussian branches, each modeling a geometrically distinct facial domain. Extensive experiments demonstrate consistently strong performance against state-of-the-art methods while achieving the fastest inference speed among all compared methods on a single GPU with the complete driving pipeline included.

📄 PDF Abstract BibTeX arXiv:2607.28032

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Operationalizing Quantized Disentanglement

2025-11-25 · Vitoria Barin-Pacela, Kartik Ahuja, Simon Lacoste-Julien, Pascal Vincent arxiv

Recent theoretical work established the unsupervised identifiability of quantized factors under any diffeomorphism. The theory assumes that quantization thresholds correspond to axis-aligned discontinuities in the probab…

SpeechSplit 2.0: Unsupervised speech disentanglement for voice conversion Without tuning autoencoder Bottlenecks

2022-03-26 · Chak Ho Chan, Kaizhi Qian, Yang Zhang, Mark Hasegawa-Johnson

SpeechSplit can perform aspect-specific voice conversion by disentangling speech into content, rhythm, pitch, and timbre using multiple autoencoders in an unsupervised manner. However, SpeechSplit requires careful tuning…

DisentanglementRhythmVoice Conversion

Faster 3D Gaussian Splatting Convergence via Structure-Aware Densification

2026-04-30 · Linjie Lyu, Ayush Tewari, Jianchun Chen, Thomas Leimkühler 외 arxiv

3D Gaussian Splatting has emerged as a powerful scene representation for real-time novel-view synthesis. However, its standard adaptive density control relies on screen-space positional gradients, which do not distinguis…

Investigation of a Data Split Strategy Involving the Time Axis in Adverse Event Prediction Using Machine Learning

2022-04-19 · Katsuhisa Morita, Tadahaya Mizuno, Hiroyuki Kusuhara

Adverse events are a serious issue in drug development and many prediction methods using machine learning have been developed. The random split cross-validation is the de facto standard for model building and evaluation …

BIG-bench Machine LearningPrediction

Real-time and Large-scale Fleet Allocation of Autonomous Taxis: A Case Study in New York Manhattan Island

2020-09-06 · Yue Yang, Wencang Bao, Mohsen Ramezani, Zhe Xu

Nowadays, autonomous taxis become a highly promising transportation mode, which helps relieve traffic congestion and avoid road accidents. However, it hinders the wide implementation of this service that traditional mode…