paper-with-me

Papers

A Self-Supervised Learning Framework for Video Encoding Complexity Clustering

2026-06-28 · Krishna Srikar Durbha, Hassene Tmar, Ping-Hao Wu, Ioannis Katsavounidis, Alan C. Bovik arxiv

Adaptive video streaming is a widely used technique for delivering video content over the internet. One of the key challenges is determining the optimal encoding settings for each video, which can vary significantly based on its content and characteristics. In this paper, we propose Compression Echo Contrastive Learning (CECL), a novel self-supervised learning framework for clustering videos based on their encoding complexity. Our method leverages the response of a video to compression - the Compression Echo - as a supervisory signal, allowing the model to capture underlying encoding characteristics during pretraining. We conduct extensive experiments to demonstrate the effectiveness of our learned representations for the downstream task of clustering videos by their encoding complexity. Our results show that CECL improves upon existing state-of-the-art visual encoders and delivers strong bitrate and quality savings against the fixed bitrate ladder.

📄 PDF Abstract BibTeX arXiv:2606.29166

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised LearningContrastive Learning

Similar Papers 제목 키워드 기반

Self-Supervised Equivariant Scene Synthesis from Video

2021-02-01 · Cinjon Resnick, Or Litany, Cosmas Heiß, Hugo Larochelle 외

We propose a self-supervised framework to learn scene representations from video that are automatically delineated into background, characters, and their animations. Our method capitalizes on moving characters being equi…

Intra Encoding Complexity Control with a Time-Cost Model for Versatile Video Coding

2022-06-13 · Yan Huang, Jizheng Xu, Li Zhang, Yan Zhao 외

For the latest video coding standard Versatile Video Coding (VVC), the encoding complexity is much higher than previous video coding standards to achieve a better coding efficiency, especially for intra coding. The compl…

Locality-Aware Inter-and Intra-Video Reconstruction for Self-Supervised Correspondence Learning

2022-03-27 · Liulei Li, Tianfei Zhou, Wenguan Wang, Lu Yang 외

Our target is to learn visual correspondence from unlabeled videos. We develop LIIR, a locality-aware inter-and intra-video reconstruction framework that fills in three missing pieces, i.e., instance discrimination, loca…

PositionRepresentation LearningVideo Reconstruction

Locality-Aware Inter- and Intra-Video Reconstruction for Self-Supervised Correspondence Learning

2022-01-01 · CVPR 2022 1 · Liulei Li, Tianfei Zhou, Wenguan Wang, Lu Yang 외

Our target is to learn visual correspondence from unlabeled videos. We develop LIIR, a locality-aware inter-and intra-video reconstruction framework that fills in three missing pieces, i.e., instance discrimination, …

PositionRepresentation LearningVideo Reconstruction

Self-Attention Based Generative Adversarial Networks For Unsupervised Video Summarization

2023-07-16 · Maria Nektaria Minaidi, Charilaos Papaioannou, Alexandros Potamianos

In this paper, we study the problem of producing a comprehensive video summary following an unsupervised approach that relies on adversarial learning. We build on a popular method where a Generative Adversarial Network (…

Generative Adversarial NetworkUnsupervised Video SummarizationVideo Summarization