paper-with-me

Papers

Compressed Video Contrastive Learning

2021-12-01 · NeurIPS 2021 12 · Yuqi Huo, Mingyu Ding, Haoyu Lu, Nanyi Fei, Zhiwu Lu, Ji-Rong Wen, Ping Luo

This work concerns self-supervised video representation learning (SSVRL), one topic that has received much attention recently. Since videos are storage-intensive and contain a rich source of visual content, models designed for SSVRL are expected to be storage- and computation-efficient, as well as effective. However, most existing methods only focus on one of the two objectives, failing to consider both at the same time. In this work, for the first time, the seemingly contradictory goals are simultaneously achieved by exploiting compressed videos and capturing mutual information between two input streams. Specifically, a novel Motion Vector based Cross Guidance Contrastive learning approach (MVCGC) is proposed. For storage and computation efficiency, we choose to directly decode RGB frames and motion vectors (that resemble low-resolution optical flows) from compressed videos on-the-fly. To enhance the representation ability of the motion vectors, hence the effectiveness of our method, we design a cross guidance contrastive learning algorithm based on multi-instance InfoNCE loss, where motion vectors can take supervision signals from RGB frames and vice versa. Comprehensive experiments on two downstream tasks show that our MVCGC yields new state-of-the-art while being significantly more efficient than its competitors.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningRepresentation Learning

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음
InfoNCE 설명 없음

Similar Papers 제목 키워드 기반

Anti-Compression Contrastive Facial Forgery Detection

2023-02-13 · Jiajun Huang, Xinqi Zhu, Chengbin Du, Siqi Ma 외

Forgery facial images and videos have increased the concern of digital security. It leads to the significant development of detecting forgery data recently. However, the data, especially the videos published on the Inter…

Contrastive Learning

End-to-End Compressed Video Representation Learning for Generic Event Boundary Detection

2022-03-29 · CVPR 2022 1 · CongCong Li, Xinyao Wang, Longyin Wen, Dexiang Hong 외

Generic event boundary detection aims to localize the generic, taxonomy-free event boundaries that segment videos into chunks. Existing methods typically require video frames to be decoded before feeding into the network…

Boundary DetectionGeneric Event Boundary DetectionRepresentation Learning

Combining Contrastive and Supervised Learning for Video Super-Resolution Detection

2022-05-20 · Viacheslav Meshchaninov, Ivan Molodetskikh, Dmitriy Vatolin

Upscaled video detection is a helpful tool in multimedia forensics, but it is a challenging task that involves various upscaling and compression algorithms. There are many resolution-enhancement methods, including interp…

Data AugmentationSuper-ResolutionVideo Super-Resolution

Hybrid Contrastive Quantization for Efficient Cross-View Video Retrieval

2022-02-07 · Jinpeng Wang, Bin Chen, Dongliang Liao, Ziyun Zeng 외

With the recent boom of video-based social platforms (e.g., YouTube and TikTok), video retrieval using sentence queries has become an important demand and attracts increasing research attention. Despite the decent perfor…

Contrastive LearningQuantizationRepresentation LearningRetrieval+2

A Diffusion Model Based Quality Enhancement Method for HEVC Compressed Video

2023-11-15 · Zheng Liu, Honggang Qi

Video post-processing methods can improve the quality of compressed videos at the decoder side. Most of the existing methods need to train corresponding models for compressed videos with different quantization parameters…

DecoderQuantization