paper-with-me

Papers

ESCA: Enabling Seamless Codec Avatar Execution through Algorithm and Hardware Co-Optimization for Virtual Reality

2025-10-27 · Mingzhi Zhu, Ding Shang, Sai Qian Zhang arxiv

Photorealistic Codec Avatars (PCA), which generate high-fidelity human face renderings, are increasingly being used in Virtual Reality (VR) environments to enable immersive communication and interaction through deep learning-based generative models. However, these models impose significant computational demands, making real-time inference challenging on resource-constrained VR devices such as head-mounted displays, where latency and power efficiency are critical. To address this challenge, we propose an efficient post-training quantization (PTQ) method tailored for Codec Avatar models, enabling low-precision execution without compromising output quality. In addition, we design a custom hardware accelerator that can be integrated into the system-on-chip of VR devices to further enhance processing efficiency. Building on these components, we introduce ESCA, a full-stack optimization framework that accelerates PCA inference on edge VR platforms. Experimental results demonstrate that ESCA boosts FovVideoVDP quality scores by up to $+0.39$ over the best 4-bit baseline, delivers up to $3.36\times$ latency reduction, and sustains a rendering rate of 100 frames per second in end-to-end tests, satisfying real-time VR requirements. These results demonstrate the feasibility of deploying high-fidelity codec avatars on resource-constrained devices, opening the door to more immersive and portable VR experiences.

📄 PDF Abstract BibTeX arXiv:2510.24787

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Efficient 3D Gaussian Human Avatar Compression: A Prior-Guided Framework

2025-10-12 · Shanzhi Yin, Bolin Chen, Xinju Wu, Ru-Ling Liao 외 arxiv

This paper proposes an efficient 3D avatar coding framework that leverages compact human priors and canonical-to-target transformation to enable high-quality 3D human avatar video compression at ultra-low bit rates. The …

Novel View SynthesisVideo Reconstruction

Pixel Codec Avatars

2021-04-09 · CVPR 2021 1 · Shugao Ma, Tomas Simon, Jason Saragih, Dawei Wang 외

Telecommunication with photorealistic avatars in virtual or augmented reality is a promising path for achieving authentic face-to-face communication in 3D over remote physical distances. In this work, we present the Pixe…

Decoder

Auto-CARD: Efficient and Robust Codec Avatar Driving for Real-time Mobile Telepresence

2023-04-24 · CVPR 2023 1 · Yonggan Fu, Yuecheng Li, Chenghui Li, Jason Saragih 외

Real-time and robust photorealistic avatars for telepresence in AR/VR have been highly desired for enabling immersive photorealistic telepresence. However, there still exists one key bottleneck: the considerable computat…

Neural Architecture Search

Large-scale Codec Avatars: The Unreasonable Effectiveness of Large-scale Avatar Pretraining

2026-04-02 · Junxuan Li, Rawal Khirodkar, Chengan He, Zhongshi Jiang 외 arxiv

High-quality 3D avatar modeling faces a critical trade-off between fidelity and generalization. On the one hand, multi-view studio data enables high-fidelity modeling of humans with precise control over expressions and p…

F-CAD: A Framework to Explore Hardware Accelerators for Codec Avatar Decoding

2021-03-08 · Xiaofan Zhang, Dawei Wang, Pierce Chuang, Shugao Ma 외

Creating virtual avatars with realistic rendering is one of the most essential and challenging tasks to provide highly immersive virtual reality (VR) experiences. It requires not only sophisticated deep neural network (D…

Decoder