paper-with-me

Papers

Resolution-Agnostic Neural Compression for High-Fidelity Portrait Video Conferencing via Implicit Radiance Fields

2024-02-26 · Yifei Li, Xiaohong Liu, Yicong Peng, Guangtao Zhai, Jun Zhou

Video conferencing has caught much more attention recently. High fidelity and low bandwidth are two major objectives of video compression for video conferencing applications. Most pioneering methods rely on classic video compression codec without high-level feature embedding and thus can not reach the extremely low bandwidth. Recent works instead employ model-based neural compression to acquire ultra-low bitrates using sparse representations of each frame such as facial landmark information, while these approaches can not maintain high fidelity due to 2D image-based warping. In this paper, we propose a novel low bandwidth neural compression approach for high-fidelity portrait video conferencing using implicit radiance fields to achieve both major objectives. We leverage dynamic neural radiance fields to reconstruct high-fidelity talking head with expression features, which are represented as frame substitution for transmission. The overall system employs deep model to encode expression features at the sender and reconstruct portrait at the receiver with volume rendering as decoder for ultra-low bandwidth. In particular, with the characteristic of neural radiance fields based model, our compression approach is resolution-agnostic, which means that the low bandwidth achieved by our approach is independent of video resolution, while maintaining fidelity for higher resolution reconstruction. Experimental results demonstrate that our novel framework can (1) construct ultra-low bandwidth video conferencing, (2) maintain high fidelity portrait and (3) have better performance on high-resolution video compression than previous works.

📄 PDF Abstract BibTeX arXiv:2402.16599

Code (0)

등록된 구현이 없습니다.

Tasks

Video Compression

Similar Papers 제목 키워드 기반

HeadsUp! High-Fidelity Portrait Image Super-Resolution

2025-10-10 · Renjie Li, Zihao Zhu, Xiaoyu Wang, Zhengzhong Tu arxiv

Portrait pictures, which typically feature both human subjects and natural backgrounds, are one of the most prevalent forms of photography on social media. Existing image super-resolution (ISR) techniques generally focus…

Image Super-Resolution

FREAK: Frequency-modulated High-fidelity and Real-time Audio-driven Talking Portrait Synthesis

2025-03-06 · Ziqi Ni, Ao Fu, Yi Zhou

Achieving high-fidelity lip-speech synchronization in audio-driven talking portrait synthesis remains challenging. While multi-stage pipelines or diffusion models yield high-quality results, they suffer from high computa…

Audio-Visual Synchronization

Enhanced Diagnostic Fidelity in Pathology Whole Slide Image Compression via Deep Learning

2025-03-14 · Maximilian Fischer, Peter Neher, Peter Schüffler, Shuhan Xiao 외

Accurate diagnosis of disease often depends on the exhaustive examination of Whole Slide Images (WSI) at microscopic resolution. Efficient handling of these data-intensive images requires lossy compression techniques. Th…

DiagnosticImage CompressionMS-SSIMSSIM+1

Hallo4: High-Fidelity Dynamic Portrait Animation via Direct Preference Optimization and Temporal Motion Modulation

2025-05-29 · Jiahao Cui, Yan Chen, Mingwang Xu, Hanlin Shang 외

Generating highly dynamic and photorealistic portrait animations driven by audio and skeletal motion remains challenging due to the need for precise lip synchronization, natural facial expressions, and high-fidelity body…

Portrait AnimationVideo Alignment

Physics-Preserving Latent Compression for Zero-Shot Resolution Transfer in 3D Turbulence

2026-06-19 · Yilong Dai, Yiming Sun, Yiheng Chen, Ziyi Wang 외 arxiv

High-resolution turbulence modeling is essential for scientific computing, but remains constrained by the cost of direct numerical simulation and the scarcity of full-resolution data. Existing scientific compressors redu…