Papers Self-supervised Video Retrieval
“Self-supervised Video Retrieval” 태그가 달린 논문 10편 · 필터 해제
Similarity Contrastive Estimation for Image and Video Soft Contrastive Self-Supervised Learning
Contrastive representation learning has proven to be an effective self-supervised learning method for images and videos. Most successful approaches are based on Noise Contrastive Estimation (NCE) and use different views …
Contrastive LearningLinear evaluationRepresentation LearningSelf-Supervised Action Recognition+4SLIC: Self-Supervised Learning with Iterative Clustering for Human Action Videos
Self-supervised methods have significantly closed the gap with end-to-end supervised learning for image classification. In the case of human action videos, however, where both appearance and motion are significant factor…
Action ClassificationClusteringContrastive Learningimage-classification+5Self-Supervised Audio-Visual Representation Learning with Relaxed Cross-Modal Synchronicity
We present CrissCross, a self-supervised framework for learning audio-visual representations. A novel notion is introduced in our framework whereby in addition to learning the intra-modal and standard 'synchronous' cross…
Audio ClassificationRetrievalSelf-Supervised Action RecognitionSelf-Supervised Audio Classification+4Self-supervised Video Representation Learning with Cross-Stream Prototypical Contrasting
Instance-level contrastive learning techniques, which rely on data augmentation and a contrastive loss function, have found great success in the domain of visual representation learning. They are not suitable for exploit…
Action RecognitionAction Recognition In VideosContrastive LearningData Augmentation+9Self-supervised Video Retrieval Transformer Network
Content-based video retrieval aims to find videos from a large video database that are similar to or even near-duplicate of a given query video. Video representation and similarity search algorithms are crucial to any vi…
RetrievalSelf-supervised Video RetrievalVideo RetrievalVideo SimilarityTCLR: Temporal Contrastive Learning for Video Representation
Contrastive learning has nearly closed the gap between supervised and self-supervised learning of image representations, and has also been explored for videos. However, prior work on contrastive learning for video data h…
Action ClassificationAction RecognitionContrastive LearningGeneral Classification+6Pretext-Contrastive Learning: Toward Good Practices in Self-supervised Video Representation Leaning
Recently, pretext-task based methods are proposed one after another in self-supervised video feature learning. Meanwhile, contrastive learning methods also yield good performance. Usually, new methods can beat previous o…
Contrastive LearningData AugmentationSelf-Supervised Action RecognitionSelf-Supervised Learning+2Self-supervised Video Representation Learning Using Inter-intra Contrastive Framework
We propose a self-supervised method to learn feature representations from videos. A standard approach in traditional self-supervised methods uses positive-negative data pairs to train with contrastive learning strategy. …
Action Recognition In VideosContrastive LearningRepresentation LearningRetrieval+4Video Playback Rate Perception for Self-Supervised Spatio-Temporal Representation Learning
In self-supervised spatio-temporal representation learning, the temporal resolution and long-short term characteristics are not yet fully explored, which limits representation capabilities of learned models. In this pape…
Action RecognitionDecoderRepresentation LearningRetrieval+3Video Cloze Procedure for Self-Supervised Spatio-Temporal Learning
We propose a novel self-supervised method, referred to as Video Cloze Procedure (VCP), to learn rich spatial-temporal representations. VCP first generates "blanks" by withholding video clips and then creates "options" by…
Action RecognitionRepresentation LearningRetrievalSelf-Supervised Action Recognition+3