AutoSSVH: Exploring Automated Frame Sampling for Efficient Self-Supervised Video Hashing
Self-Supervised Video Hashing (SSVH) compresses videos into hash codes for efficient indexing and retrieval using unlabeled training videos. Existing approaches rely on random frame sampling to learn video features and treat all frames equally. This results in suboptimal hash codes, as it ignores frame-specific information density and reconstruction difficulty. To address this limitation, we propose a new framework, termed AutoSSVH, that employs adversarial frame sampling with hash-based contrastive learning. Our adversarial sampling strategy automatically identifies and selects challenging frames with richer information for reconstruction, enhancing encoding capability. Additionally, we introduce a hash component voting strategy and a point-to-set (P2Set) hash-based contrastive objective, which help capture complex inter-video semantic relationships in the Hamming space and improve the discriminability of learned hash codes. Extensive experiments demonstrate that AutoSSVH achieves superior retrieval efficacy and efficiency compared to state-of-the-art approaches. Code is available at https://github.com/EliSpectre/CVPR25-AutoSSVH.
Code (1)
Tasks
Contrastive LearningRetrievalSimilar Papers 제목 키워드 기반
Self-Supervised Time Series Representation Learning by Inter-Intra Relational Reasoning
Self-supervised learning achieves superior performance in many domains by extracting useful representations from the unlabeled data. However, most of traditional self-supervised methods mainly focus on exploring the inte…
RelationRelational ReasoningRepresentation LearningSelf-Supervised Learning+3Exploring the Effectiveness of Using LLMs for Automated Assessment of Student Self Explanations in Programming Education
Worked examples are step-by-step solutions to problems in a specific domain, offered to students to acquire domain-specific problem-solving skills. The effectiveness of worked examples could be enhanced by combining them…
Binary ClassificationSemantic SimilarityToward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning
Tool-Integrated Reasoning (TIR) enables large language models (LLMs) to improve their internal reasoning ability by integrating external tools. However, models employing TIR often display suboptimal behaviors, such as in…
Exploring Relations in Untrimmed Videos for Self-Supervised Learning
Existing video self-supervised learning methods mainly rely on trimmed videos for model training. However, trimmed datasets are manually annotated from untrimmed videos. In this sense, these methods are not really self-s…
Action RecognitionChange DetectionRetrievalSelf-Supervised Learning+1Augmenting Automated Game Testing with Deep Reinforcement Learning
General game testing relies on the use of human play testers, play test scripting, and prior knowledge of areas of interest to produce relevant test data. Using deep reinforcement learning (DRL), we introduce a self-lear…
Deep Reinforcement LearningFPS Gamesreinforcement-learningReinforcement Learning+2