paper-with-me

홈 › Papers

Tactile-JEPA: Topology-Aware Self-Supervised Representation Learning for Distributed Tactile Sensors

2026-09-21 · Elizaveta Kovtun, Matvey Konovalov, Andrey Sakhovskiy, Semen Budennyy hf

Tactile sensing is an essential modality for robots performing contact-rich, dexterous manipulation, particularly under visual occlusion. While pre-trained image encoders are standard in robot learning pipelines, tactile encoders are still commonly trained from scratch from raw, noisy signals, which might limit their expressivity. Existing self-supervised learning (SSL) approaches focus predominantly on vision-based tactile sensors, leaving distributed electronic skins largely unaddressed. These sensors, however, have a distinctive property: their sensing elements are sparse and irregularly arranged over the surface they cover, which makes direct reuse of visual SSL methods suboptimal. We present Tactile-JEPA, an efficient self-supervised pre-training method that uses the spatial arrangement of tactile sensors to learn topology-aware representations. Specifically, it is trained to predict the embeddings of masked sensing elements from the unmasked remainder, using the sensor connectivity graph to guide spatial masking. Our analysis shows that effective tactile representations require capturing both local contact details and the global state of the tactile surface, which we achieve through dual-scale masking. Across three diverse datasets spanning magnetic and piezoresistive sensors, different robot embodiments, and single- and paired-sensor configurations, Tactile-JEPA reduces force estimation error by 6.3% and in-hand orientation error by 20.8% over the prior state-of-the-art, with consistent gains in other downstream applications, including policy learning. Overall, our results demonstrate that the benefit of tactile sensing depends critically on the quality of encoder pre-training, a problem which Tactile-JEPA addresses directly. Code is available at https://github.com/E-Kovtun/tactile.

📄 PDF Abstract BibTeX arXiv:2609.24385

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised LearningRepresentation Learning

Similar Papers 제목 키워드 기반

Entity-Centric World Models: Interaction-Aware Masking for Causal Video Prediction

2026-05-14 · Santosh Kumar Paidi arxiv

Learning predictive world models from unlabelled video is a foundational challenge in artificial intelligence. While Joint Embedding Predictive Architectures (JEPA) have set new benchmarks in semantic classification, the…

Video Prediction

3D-JEPA: A Joint Embedding Predictive Architecture for 3D Self-Supervised Representation Learning

2024-09-24 · Naiwen Hu, Haozhe Cheng, Yifan Xie, Shiqi Li 외

Invariance-based and generative methods have shown a conspicuous performance for 3D self-supervised representation learning (SSRL). However, the former relies on hand-crafted data augmentations that introduce bias not un…

3D Part Segmentation3D Point Cloud ClassificationDecoderFew-Shot 3D Point Cloud Classification+1

KerJEPA: Kernel Discrepancies for Euclidean Self-Supervised Learning

2025-12-22 · Eric Zimmermann, Harley Wiltzer, Justin Szeto, David Alvarez-Melis 외 arxiv

Recent breakthroughs in self-supervised Joint-Embedding Predictive Architectures (JEPAs) have established that regularizing Euclidean representations toward isotropic Gaussian priors yields provable gains in training sta…

Self-Supervised Learning

JEPADepth: Masked Predictive Representation Learning for Self-Supervised Monocular Depth Estimation

2026-07-29 · Ionuţ Grigore, Călin-Adrian Popa arxiv

Self-supervised monocular depth estimation typically relies on photometric reconstruction losses that couple depth, pose, and appearance assumptions. In this paper, we propose JEPADepth, a self-supervised monocular depth…

Monocular Depth EstimationRepresentation Learning

AD-L-JEPA: Self-Supervised Spatial World Models with Joint Embedding Predictive Architecture for Autonomous Driving with LiDAR Data

2025-01-09 · Haoran Zhu, Zhenyuan Dong, Kristi Topollai, Anna Choromanska

As opposed to human drivers, current autonomous driving systems still require vast amounts of labeled data to train. Recently, world models have been proposed to simultaneously enhance autonomous driving capabilities by …

3D Object DetectionAutonomous DrivingContrastive LearningTransfer Learning