paper-with-me

Papers Self-supervised Video Retrieval

“Self-supervised Video Retrieval” 태그가 달린 논문 10편 · 필터 해제

Similarity Contrastive Estimation for Image and Video Soft Contrastive Self-Supervised Learning

2022-12-21 · Julien Denize, Jaonary Rabarisoa, Astrid Orcesi, Romain Hérault

Contrastive representation learning has proven to be an effective self-supervised learning method for images and videos. Most successful approaches are based on Noise Contrastive Estimation (NCE) and use different views …

Contrastive LearningLinear evaluationRepresentation LearningSelf-Supervised Action Recognition+4

SLIC: Self-Supervised Learning with Iterative Clustering for Human Action Videos

2022-06-25 · CVPR 2022 1 · Salar Hosseini Khorasgani, Yuxuan Chen, Florian Shkurti

Self-supervised methods have significantly closed the gap with end-to-end supervised learning for image classification. In the case of human action videos, however, where both appearance and motion are significant factor…

Action ClassificationClusteringContrastive Learningimage-classification+5

Self-Supervised Audio-Visual Representation Learning with Relaxed Cross-Modal Synchronicity

2021-11-09 · Pritam Sarkar, Ali Etemad

We present CrissCross, a self-supervised framework for learning audio-visual representations. A novel notion is introduced in our framework whereby in addition to learning the intra-modal and standard 'synchronous' cross…

Audio ClassificationRetrievalSelf-Supervised Action RecognitionSelf-Supervised Audio Classification+4

Self-supervised Video Representation Learning with Cross-Stream Prototypical Contrasting

2021-06-18 · Martine Toering, Ioannis Gatopoulos, Maarten Stol, Vincent Tao Hu

Instance-level contrastive learning techniques, which rely on data augmentation and a contrastive loss function, have found great success in the domain of visual representation learning. They are not suitable for exploit…

Action RecognitionAction Recognition In VideosContrastive LearningData Augmentation+9

Self-supervised Video Retrieval Transformer Network

2021-04-16 · Xiangteng He, Yulin Pan, Mingqian Tang, Yiliang Lv

Content-based video retrieval aims to find videos from a large video database that are similar to or even near-duplicate of a given query video. Video representation and similarity search algorithms are crucial to any vi…

RetrievalSelf-supervised Video RetrievalVideo RetrievalVideo Similarity

TCLR: Temporal Contrastive Learning for Video Representation

2021-01-20 · Ishan Dave, Rohit Gupta, Mamshad Nayeem Rizve, Mubarak Shah

Contrastive learning has nearly closed the gap between supervised and self-supervised learning of image representations, and has also been explored for videos. However, prior work on contrastive learning for video data h…

Action ClassificationAction RecognitionContrastive LearningGeneral Classification+6

Pretext-Contrastive Learning: Toward Good Practices in Self-supervised Video Representation Leaning

2020-10-29 · Li Tao, Xueting Wang, Toshihiko Yamasaki

Recently, pretext-task based methods are proposed one after another in self-supervised video feature learning. Meanwhile, contrastive learning methods also yield good performance. Usually, new methods can beat previous o…

Contrastive LearningData AugmentationSelf-Supervised Action RecognitionSelf-Supervised Learning+2

Self-supervised Video Representation Learning Using Inter-intra Contrastive Framework

2020-08-06 · Li Tao, Xueting Wang, Toshihiko Yamasaki

We propose a self-supervised method to learn feature representations from videos. A standard approach in traditional self-supervised methods uses positive-negative data pairs to train with contrastive learning strategy. …

Action Recognition In VideosContrastive LearningRepresentation LearningRetrieval+4

Video Playback Rate Perception for Self-Supervised Spatio-Temporal Representation Learning

2020-06-01 · CVPR 2020 6 · Yuan Yao, Chang Liu, Dezhao Luo, Yu Zhou 외

In self-supervised spatio-temporal representation learning, the temporal resolution and long-short term characteristics are not yet fully explored, which limits representation capabilities of learned models. In this pape…

Action RecognitionDecoderRepresentation LearningRetrieval+3

Video Cloze Procedure for Self-Supervised Spatio-Temporal Learning

2020-01-02 · Dezhao Luo, Chang Liu, Yu Zhou, Dongbao Yang 외

We propose a novel self-supervised method, referred to as Video Cloze Procedure (VCP), to learn rich spatial-temporal representations. VCP first generates "blanks" by withholding video clips and then creates "options" by…

Action RecognitionRepresentation LearningRetrievalSelf-Supervised Action Recognition+3
1–10 / 10