paper-with-me

홈 › Papers

Generative Face Video Coding Techniques and Standardization Efforts: A Review

2023-11-05 · Bolin Chen, Jie Chen, Shiqi Wang, Yan Ye

Generative Face Video Coding (GFVC) techniques can exploit the compact representation of facial priors and the strong inference capability of deep generative models, achieving high-quality face video communication in ultra-low bandwidth scenarios. This paper conducts a comprehensive survey on the recent advances of the GFVC techniques and standardization efforts, which could be applicable to ultra low bitrate communication, user-specified animation/filtering and metaverse-related functionalities. In particular, we generalize GFVC systems within one coding framework and summarize different GFVC algorithms with their corresponding visual representations. Moreover, we review the GFVC standardization activities that are specified with supplemental enhancement information messages. Finally, we discuss fundamental challenges and broad applications on GFVC techniques and their standardization potentials, as well as envision their future trends. The project page can be found at https://github.com/Berlin0610/Awesome-Generative-Face-Video-Coding.

📄 PDF Abstract BibTeX arXiv:2311.02649

Code (1)

berlin0610/awesome-generative-face-video-coding 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Generative Models at the Frontier of Compression: A Survey on Generative Face Video Coding

2025-06-09 · Bolin Chen, Shanzhi Yin, Goluck Konuko, Giuseppe Valenzise 외

The rise of deep generative models has greatly advanced video compression, reshaping the paradigm of face video coding through their powerful capability for semantic-aware representation and lifelike synthesis. Generativ…

BenchmarkingVideo Compression

Standardizing Generative Face Video Compression using Supplemental Enhancement Information

2024-10-19 · Bolin Chen, Yan Ye, Jie Chen, Ru-Ling Liao 외

This paper proposes a Generative Face Video Compression (GFVC) approach using Supplemental Enhancement Information (SEI), where a series of compact spatial and temporal representations of a face video signal (i.e., 2D/3D…

Video Compression

Towards Coding for Human and Machine Vision: A Scalable Image Coding Approach

2020-01-09 · Yueyu Hu, Shuai Yang, Wenhan Yang, Ling-Yu Duan 외

The past decades have witnessed the rapid development of image and video coding techniques in the era of big data. However, the signal fidelity-driven coding pipeline design limits the capability of the existing image/vi…

Facial Landmark DetectionImage Reconstruction

AI Oriented Large-Scale Video Management for Smart City: Technologies, Standards and Beyond

2017-12-05 · Ling-Yu Duan, Yihang Lou, Shiqi Wang, Wen Gao 외

Deep learning has achieved substantial success in a series of tasks in computer vision. Intelligent video analysis, which can be broadly applied to video surveillance in various smart city applications, can also be drive…

Deep LearningManagement

An Emerging Coding Paradigm VCM: A Scalable Coding Approach Beyond Feature and Signal

2020-01-09 · Sifeng Xia, Kunchangtai Liang, Wenhan Yang, Ling-Yu Duan 외

In this paper, we study a new problem arising from the emerging MPEG standardization effort Video Coding for Machine (VCM), which aims to bridge the gap between visual feature compression and classical video coding. VCM …

Action RecognitionFeature CompressionSSIM