paper-with-me

Papers

LetsMap: Unsupervised Representation Learning for Semantic BEV Mapping

2024-05-29 · Nikhil Gosala, Kürsat Petek, B Ravi Kiran, Senthil Yogamani, Paulo Drews-Jr, Wolfram Burgard, Abhinav Valada

Semantic Bird's Eye View (BEV) maps offer a rich representation with strong occlusion reasoning for various decision making tasks in autonomous driving. However, most BEV mapping approaches employ a fully supervised learning paradigm that relies on large amounts of human-annotated BEV ground truth data. In this work, we address this limitation by proposing the first unsupervised representation learning approach to generate semantic BEV maps from a monocular frontal view (FV) image in a label-efficient manner. Our approach pretrains the network to independently reason about scene geometry and scene semantics using two disjoint neural pathways in an unsupervised manner and then finetunes it for the task of semantic BEV mapping using only a small fraction of labels in the BEV. We achieve label-free pretraining by exploiting spatial and temporal consistency of FV images to learn scene geometry while relying on a novel temporal masked autoencoder formulation to encode the scene representation. Extensive evaluations on the KITTI-360 and nuScenes datasets demonstrate that our approach performs on par with the existing state-of-the-art approaches while using only 1% of BEV labels and no additional labeled data.

📄 PDF Abstract BibTeX arXiv:2405.18852

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingDecision MakingRepresentation Learning

Similar Papers 제목 키워드 기반

Cross-topic distributional semantic representations via unsupervised mappings

2019-04-11 · NAACL 2019 6 · Eleftheria Briakou, Nikos Athanasiou, Alexandros Potamianos

In traditional Distributional Semantic Models (DSMs) the multiple senses of a polysemous word are conflated into a single vector space representation. In this work, we propose a DSM that learns multiple distributional re…

Word Similarity

Adversarial Feature Learning

2016-05-31 · Jeff Donahue, Philipp Krähenbühl, Trevor Darrell

The ability of the Generative Adversarial Networks (GANs) framework to learn generative models mapping from simple latent distributions to arbitrarily complex data distributions has been demonstrated empirically, with co…

Unsupervised Domain Generalization by Learning a Bridge Across Domains

2021-12-04 · CVPR 2022 1 · Sivan Harary, Eli Schwartz, Assaf Arbelle, Peter Staar 외

The ability to generalize learned representations across significantly different visual domains, such as between real photos, clipart, paintings, and sketches, is a fundamental capacity of the human visual system. In thi…

Domain GeneralizationSelf-Supervised Learning

XGAN: Unsupervised Image-to-Image Translation for Many-to-Many Mappings

2017-11-14 · ICLR 2018 1 · Amélie Royer, Konstantinos Bousmalis, Stephan Gouws, Fred Bertsch 외

Style transfer usually refers to the task of applying color and texture information from a specific style image to a given content image while preserving the structure of the latter. Here we tackle the more generic probl…

Domain AdaptationImage-to-Image TranslationStyle TransferTranslation+1

Incremental Semantic Mapping with Unsupervised On-line Learning

2019-07-09 · Ygor C. N. Sousa, Hansenclever F. Bassani

This paper introduces an incremental semantic mapping approach, with on-line unsupervised learning, based on Self-Organizing Maps (SOM) for robotic agents. The method includes a mapping module, which incrementally create…

Clustering