paper-with-me

홈 › Papers

Connecting Joint-Embedding Predictive Architecture with Contrastive Self-supervised Learning

2024-10-25 · Shentong Mo, Shengbang Tong

In recent advancements in unsupervised visual representation learning, the Joint-Embedding Predictive Architecture (JEPA) has emerged as a significant method for extracting visual features from unlabeled imagery through an innovative masking strategy. Despite its success, two primary limitations have been identified: the inefficacy of Exponential Moving Average (EMA) from I-JEPA in preventing entire collapse and the inadequacy of I-JEPA prediction in accurately learning the mean of patch representations. Addressing these challenges, this study introduces a novel framework, namely C-JEPA (Contrastive-JEPA), which integrates the Image-based Joint-Embedding Predictive Architecture with the Variance-Invariance-Covariance Regularization (VICReg) strategy. This integration is designed to effectively learn the variance/covariance for preventing entire collapse and ensuring invariance in the mean of augmented views, thereby overcoming the identified limitations. Through empirical and theoretical evaluations, our work demonstrates that C-JEPA significantly enhances the stability and quality of visual representation learning. When pre-trained on the ImageNet-1K dataset, C-JEPA exhibits rapid and improved convergence in both linear probing and fine-tuning performance metrics.

📄 PDF Abstract BibTeX arXiv:2410.19560

Code (0)

등록된 구현이 없습니다.

Tasks

Representation LearningSelf-Supervised Learning

Similar Papers 제목 키워드 기반

MJEPA: A Simple and Scalable Joint-Embedding Predictive Architecture for Audio-Visual Learning

2026-06-23 · Revant Teotia, Adrien Bardes, Michael Rabbat, Sumit Chopra 외 arxiv

Self-supervised learning from large-scale video data has emerged as a dominant paradigm for visual representation learning. Since audio and visual streams naturally co-occur in video data, extending this success to joint…

Self-Supervised LearningRepresentation Learning

Leveraging Joint Predictive Embedding and Bayesian Inference in Graph Self Supervised Learning

2025-02-02 · Srinitish Srinivasan, Omkumar CU

Graph representation learning has emerged as a cornerstone for tasks like node classification and link prediction, yet prevailing self-supervised learning (SSL) methods face challenges such as computational inefficiency,…

Bayesian InferenceGraph Representation LearningLink PredictionNode Classification+3

AD-L-JEPA: Self-Supervised Spatial World Models with Joint Embedding Predictive Architecture for Autonomous Driving with LiDAR Data

2025-01-09 · Haoran Zhu, Zhenyuan Dong, Kristi Topollai, Anna Choromanska

As opposed to human drivers, current autonomous driving systems still require vast amounts of labeled data to train. Recently, world models have been proposed to simultaneously enhance autonomous driving capabilities by …

3D Object DetectionAutonomous DrivingContrastive LearningTransfer Learning

Zero-shot Musical Stem Retrieval with Joint-Embedding Predictive Architectures

2024-11-29 · Alain Riou, Antonin Gagneré, Gaëtan Hadjeres, Stefan Lattner 외

In this paper, we tackle the task of musical stem retrieval. Given a musical mix, it consists in retrieving a stem that would fit with it, i.e., that would sound pleasant if played together. To do so, we introduce a new …

Beat TrackingContrastive LearningRetrieval

Graph-level Representation Learning with Joint-Embedding Predictive Architectures

2023-09-27 · Geri Skenderi, Hang Li, Jiliang Tang, Marco Cristani

Joint-Embedding Predictive Architectures (JEPAs) have recently emerged as a novel and powerful technique for self-supervised representation learning. They aim to learn an energy-based model by predicting the latent repre…

Contrastive LearningData AugmentationGraph ClassificationGraph Regression+1