paper-with-me

Papers

Towards Hybrid-Optimization Video Coding

2022-07-12 · Shuai Huo, Dong Liu, Li Li, Siwei Ma, Feng Wu, Wen Gao

Video coding is a mathematical optimization problem of rate and distortion essentially. To solve this complex optimization problem, two popular video coding frameworks have been developed: block-based hybrid video coding and end-to-end learned video coding. If we rethink video coding from the perspective of optimization, we find that the existing two frameworks represent two directions of optimization solutions. Block-based hybrid coding represents the discrete optimization solution because those irrelevant coding modes are discrete in mathematics. It searches for the best one among multiple starting points (i.e. modes). However, the search is not efficient enough. On the other hand, end-to-end learned coding represents the continuous optimization solution because the gradient descent is based on a continuous function. It optimizes a group of model parameters efficiently by the numerical algorithm. However, limited by only one starting point, it is easy to fall into the local optimum. To better solve the optimization problem, we propose to regard video coding as a hybrid of the discrete and continuous optimization problem, and use both search and numerical algorithm to solve it. Our idea is to provide multiple discrete starting points in the global space and optimize the local optimum around each point by numerical algorithm efficiently. Finally, we search for the global optimum among those local optimums. Guided by the hybrid optimization idea, we design a hybrid optimization video coding framework, which is built on continuous deep networks entirely and also contains some discrete modes. We conduct a comprehensive set of experiments. Compared to the continuous optimization framework, our method outperforms pure learned video coding methods. Meanwhile, compared to the discrete optimization framework, our method achieves comparable performance to HEVC reference software HM16.10 in PSNR.

📄 PDF Abstract BibTeX arXiv:2207.05565

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Task-driven Semantic Coding via Reinforcement Learning

2021-06-07 · Xin Li, Jun Shi, Zhibo Chen

Task-driven semantic video/image coding has drawn considerable attention with the development of intelligent media applications, such as license plate detection, face detection, and medical diagnosis, which focuses on ma…

Face DetectionLicense Plate DetectionMedical DiagnosisQuantization+3

Key frames assisted hybrid encoding for photorealistic compressive video sensing

2022-07-26 · Honghao Huang, Jiajie Teng, Yu Liang, Chengyang Hu 외

Snapshot compressive imaging (SCI) encodes high-speed scene video into a snapshot measurement and then computationally makes reconstructions, allowing for efficient high-dimensional data acquisition. Numerous algorithms,…

Optical Flow Estimation

Generative Compression for Face Video: A Hybrid Scheme

2022-04-21 · Anni Tang, Yan Huang, Jun Ling, ZhiYu Zhang 외

As the latest video coding standard, versatile video coding (VVC) has shown its ability in retaining pixel quality. To excavate more compression potential for video conference scenarios under ultra-low bitrate, this pape…

BVI-AOM: A New Training Dataset for Deep Video Compression Optimization

2024-08-06 · Jakub Nawała, YuXuan Jiang, Fan Zhang, Xiaoqing Zhu 외

Deep learning is now playing an important role in enhancing the performance of conventional hybrid video codecs. These learning-based methods typically require diverse and representative training material for optimizatio…

Video Compression

CANF-VC: Conditional Augmented Normalizing Flows for Video Compression

2022-07-12 · Yung-Han Ho, Chih-Peng Chang, Peng-Yu Chen, Alessandro Gnutti 외

This paper presents an end-to-end learning-based video compression system, termed CANF-VC, based on conditional augmented normalizing flows (CANF). Most learned video compression systems adopt the same hybrid-based codin…

Video Compression