paper-with-me

Papers

Mine-JEPA: In-Domain Self-Supervised Learning for Mine-Like Object Classification in Side-Scan Sonar

2026-04-01 · Taeyoun Kwon, Youngwon Choi, Hyeonyu Kim, Myeongkyun Cho, Junhyeok Choi, Moon Hwan Kim arxiv

Side-scan sonar (SSS) mine classification is a challenging maritime vision problem characterized by extreme data scarcity and a large domain gap from natural images. While self-supervised learning (SSL) and general-purpose vision foundation models have shown strong performance in general vision and several specialized domains, their use in SSS remains largely unexplored. We present Mine-JEPA, the first in-domain SSL pipeline for SSS mine classification, using SIGReg, a regularization-based SSL loss, to pretrain on only 1,170 unlabeled sonar images. In the binary mine vs. non-mine setting, Mine-JEPA achieves an F1 score of 0.935, outperforming fine-tuned DINOv3 (0.922), a foundation model pretrained on 1.7B images. For 3-class mine-like object classification, Mine-JEPA reaches 0.820 with synthetic data augmentation, again outperforming fine-tuned DINOv3 (0.810). We further observe that applying in-domain SSL to foundation models degrades performance by 10--13 percentage points, suggesting that stronger pretrained models do not always benefit from additional domain adaptation. In addition, Mine-JEPA with a compact ViT-Tiny backbone achieves competitive performance while using 4x fewer parameters than DINOv3. These results suggest that carefully designed in-domain self-supervised learning is a viable alternative to much larger foundation models in data-scarce maritime sonar imagery.

📄 PDF Abstract BibTeX arXiv:2604.00383

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised LearningData AugmentationDomain Adaptation

Similar Papers 제목 키워드 기반

Masked and Predictive Self-Supervised Foundation Models for 3D Brain MRI

2026-06-11 · Esra Ergün, Hersh Chandarana, Dan Sodickson, Gözde Ünal arxiv

Self-supervised foundation models have shown strong promise in medical imaging. However, existing MRI foundation-model studies have primarily emphasized segmentation and dense prediction tasks, while systematic investiga…

Representation Learning

BERT-JEPA: Reorganizing CLS Embeddings for Language-Invariant Semantics

2026-01-01 · Taj Gillin, Adam Lalani, Kenneth Zhang, Marcel Mateos Salles arxiv

Joint Embedding Predictive Architectures (JEPA) are a novel self supervised training technique that have shown recent promise across domains. We introduce BERT-JEPA (BEPA), a training paradigm that adds a JEPA training o…

T-SAR-JEPA: Self-Supervised Temporal Anomaly Detection in SAR Amplitude Stacks via Latent Prediction

2026-06-04 · Kerod Woldesenbet, Abem Woldesenbet arxiv

We present T-SAR-JEPA, a self-supervised framework for temporal anomaly detection in SAR amplitude stacks via latent prediction. A ViT-Base/16 encoder from SAR-JEPA is domain-adapted on 39,300 Capella patches using local…

Anomaly Detection

LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics

2025-11-11 · Randall Balestriero, Yann LeCun arxiv

Learning manipulable representations of the world and its dynamics is central to AI. Joint-Embedding Predictive Architectures (JEPAs) offer a promising blueprint, but lack of practical guidance and theory has led to ad-h…

Self-Supervised Learning

REJEPA: A Novel Joint-Embedding Predictive Architecture for Efficient Remote Sensing Image Retrieval

2025-04-04 · Shabnam Choudhury, Yash Salunkhe, Sarthak Mehrotra, Biplab Banerjee

The rapid expansion of remote sensing image archives demands the development of strong and efficient techniques for content-based image retrieval (RS-CBIR). This paper presents REJEPA (Retrieval with Joint-Embedding Pred…

Computational EfficiencyContent-Based Image RetrievalImage RetrievalRetrieval