paper-with-me

Papers

CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping

2023-10-11 · Tim Lebailly, Thomas Stegmüller, Behzad Bozorgtabar, Jean-Philippe Thiran, Tinne Tuytelaars

Leveraging nearest neighbor retrieval for self-supervised representation learning has proven beneficial with object-centric images. However, this approach faces limitations when applied to scene-centric datasets, where multiple objects within an image are only implicitly captured in the global representation. Such global bootstrapping can lead to undesirable entanglement of object representations. Furthermore, even object-centric datasets stand to benefit from a finer-grained bootstrapping approach. In response to these challenges, we introduce a novel Cross-Image Object-Level Bootstrapping method tailored to enhance dense visual representation learning. By employing object-level nearest neighbor bootstrapping throughout the training, CrIBo emerges as a notably strong and adequate candidate for in-context learning, leveraging nearest neighbor retrieval at test time. CrIBo shows state-of-the-art performance on the latter task while being highly competitive in more standard downstream segmentation tasks. Our code and pretrained models are publicly available at https://github.com/tileb1/CrIBo.

📄 PDF Abstract BibTeX arXiv:2310.07855

Code (1)

tileb1/cribo 공식 구현 pytorch

Tasks

In-Context LearningObjectRepresentation LearningRetrievalSelf-Supervised Learning

Similar Papers 제목 키워드 기반

S3PT: Scene Semantics and Structure Guided Clustering to Boost Self-Supervised Pre-Training for Autonomous Driving

2024-10-30 · Maciej K. Wozniak, Hariprasath Govindarajan, Marvin Klingner, Camille Maurice 외

Recent self-supervised clustering-based pre-training techniques like DINO and Cribo have shown impressive results for downstream detection and segmentation tasks. However, real-world applications such as autonomous drivi…

3D Object DetectionAutonomous DrivingClusteringDiversity+4

Self-supervised Learning from a Multi-view Perspective

2020-06-10 · ICLR 2021 1 · Yao-Hung Hubert Tsai, Yue Wu, Ruslan Salakhutdinov, Louis-Philippe Morency

As a subset of unsupervised representation learning, self-supervised representation learning adopts self-defined signals as supervision and uses the learned representation for downstream tasks, such as object detection a…

Image CaptioningLanguage Modellingobject-detectionObject Detection+2

Scriboora: Rethinking Human Pose Forecasting

2025-11-19 · Daniel Bermuth, Alexander Poeppel, Wolfgang Reif arxiv

Human pose forecasting predicts future poses based on past observations, and has many significant applications in areas such as action recognition, autonomous driving or human-robot interaction. This paper evaluates a wi…

Human Pose ForecastingAction RecognitionAutonomous DrivingPose Estimation

Self-supervised Feature Learning by Cross-modality and Cross-view Correspondences

2020-04-13 · Longlong Jing, Yu-cheng Chen, Ling Zhang, Mingyi He 외

The success of supervised learning requires large-scale ground truth labels which are very expensive, time-consuming, or may need special skills to annotate. To address this issue, many self- or un-supervised methods are…

3D Part Segmentation3D Shape Classification3D Shape Recognition3D Shape Retrieval+2

SCOPS: Self-Supervised Co-Part Segmentation

2019-05-03 · CVPR 2019 6 · Wei-Chih Hung, Varun Jampani, Sifei Liu, Pavlo Molchanov 외

Parts provide a good intermediate representation of objects that is robust with respect to the camera, pose and appearance variations. Existing works on part segmentation is dominated by supervised approaches that rely o…

ObjectSegmentationUnsupervised Facial Landmark DetectionUnsupervised Human Pose Estimation+2