paper-with-me

Papers

Bootstrapped Self-Supervised Training with Monocular Video for Semantic Segmentation and Depth Estimation

2021-03-19 · Yihao Zhang, John J. Leonard

For a robot deployed in the world, it is desirable to have the ability of autonomous learning to improve its initial pre-set knowledge. We formalize this as a bootstrapped self-supervised learning problem where a system is initially bootstrapped with supervised training on a labeled dataset and we look for a self-supervised training method that can subsequently improve the system over the supervised training baseline using only unlabeled data. In this work, we leverage temporal consistency between frames in monocular video to perform this bootstrapped self-supervised training. We show that a well-trained state-of-the-art semantic segmentation network can be further improved through our method. In addition, we show that the bootstrapped self-supervised training framework can help a network learn depth estimation better than pure supervised training or self-supervised training.

📄 PDF Abstract BibTeX arXiv:2103.11031

Code (0)

등록된 구현이 없습니다.

Tasks

Depth EstimationSelf-Supervised LearningSemantic Segmentation

Similar Papers 제목 키워드 기반

Self-supervised Dense 3D Reconstruction from Monocular Endoscopic Video

2019-09-06 · Xingtong Liu, Ayushi Sinha, Masaru Ishii, Gregory D. Hager 외

We present a self-supervised learning-based pipeline for dense 3D reconstruction from full-length monocular endoscopic videos without a priori modeling of anatomy or shading. Our method only relies on unlabeled monocular…

3D ReconstructionAnatomySelf-Supervised Learning

Geometric Reciprocity: Unlocking Self-Supervision for Stereoscopic Video Generation

2026-07-06 · Jingyi Lu, Kai Han arxiv

Monocular-to-stereo conversion synthesizes stereoscopic content from 2D videos for immersive 3D experiences. In modern Depth-Image-Based Rendering (DIBR) approaches, stereo inpainting of disocclusions is the critical bot…

Self-Supervised LearningVideo Generation

NimbleD: Enhancing Self-supervised Monocular Depth Estimation with Pseudo-labels and Large-scale Video Pre-training

2024-08-26 · Albert Luginov, Muhammad Shahzad

We introduce NimbleD, an efficient self-supervised monocular depth estimation learning framework that incorporates supervision from pseudo-labels generated by a large vision model. This framework does not require camera …

Depth EstimationMonocular Depth Estimation

Synthesizing Light Field Video from Monocular Video

2022-07-21 · Shrisudhan Govindarajan, Prasan Shedligeri, Sarah, Kaushik Mitra

The hardware challenges associated with light-field(LF) imaging has made it difficult for consumers to access its benefits like applications in post-capture focus and aperture control. Learning-based techniques which sol…

Self-Supervised LearningVideo Reconstruction

$S^3$Net: Semantic-Aware Self-supervised Depth Estimation with Monocular Videos and Synthetic Data

2020-07-28 · Bin Cheng, Inderjot Singh Saggu, Raunak Shah, Gaurav Bansal 외

Solving depth estimation with monocular cameras enables the possibility of widespread use of cameras as low-cost depth estimation sensors in applications such as autonomous driving and robotics. However, learning such a …

Autonomous DrivingDepth EstimationDomain AdaptationPanoptic Segmentation