paper-with-me

홈 › Papers

LatentBKI: Open-Dictionary Continuous Mapping in Visual-Language Latent Spaces with Quantifiable Uncertainty

2024-10-15 · Joey Wilson, Ruihan Xu, Yile Sun, Parker Ewen, Minghan Zhu, Kira Barton, Maani Ghaffari

This paper introduces a novel probabilistic mapping algorithm, LatentBKI, which enables open-vocabulary mapping with quantifiable uncertainty. Traditionally, semantic mapping algorithms focus on a fixed set of semantic categories which limits their applicability for complex robotic tasks. Vision-Language (VL) models have recently emerged as a technique to jointly model language and visual features in a latent space, enabling semantic recognition beyond a predefined, fixed set of semantic classes. LatentBKI recurrently incorporates neural embeddings from VL models into a voxel map with quantifiable uncertainty, leveraging the spatial correlations of nearby observations through Bayesian Kernel Inference (BKI). LatentBKI is evaluated against similar explicit semantic mapping and VL mapping frameworks on the popular Matterport3D and Semantic KITTI datasets, demonstrating that LatentBKI maintains the probabilistic benefits of continuous mapping with the additional benefit of open-dictionary queries. Real-world experiments demonstrate applicability to challenging indoor environments.

📄 PDF Abstract BibTeX arXiv:2410.11783

Code (1)

UMich-CURLY/LatentBKI 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

Focus 설명 없음
SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Word Embedding Visualization Via Dictionary Learning

2019-10-09 · Juexiao Zhang, Yubei Chen, Brian Cheung, Bruno A. Olshausen

Co-occurrence statistics based word embedding techniques have proved to be very useful in extracting the semantic and syntactic representation of words as low dimensional continuous vectors. In this work, we discovered t…

Dictionary Learning

Group Sparse Coding

2009-12-01 · NeurIPS 2009 12 · Samy Bengio, Fernando Pereira, Yoram Singer, Dennis Strelow

Bag-of-words document representations are often used in text, image and video processing. While it is relatively easy to determine a suitable word dictionary for text documents, there is no simple mapping from raw images…

Computational EfficiencyGeneral Classificationimage-classificationImage Classification+1

The ACoLi Dictionary Graph

2020-05-01 · LREC 2020 5 · Christian Chiarcos, Christian F{\"a}th, Maxim Ionov

In this paper, we report the release of the ACoLi Dictionary Graph, a large-scale collection of multilingual open source dictionaries available in two machine-readable formats, a graph representation in RDF, using the On…

Translation

Mapping and Generating Classifiers using an Open Chinese Ontology

2016-01-01 · GWC 2016 1 · Luis Morgado Da Costa, Francis Bond, Helena Gao

In languages such as Chinese, classifiers (CLs) play a central role in the quantification of noun-phrases. This can be a problem when generating text from input that does not specify the classifier, as in machine transla…

Machine TranslationTranslation

A Bayesian Approach to Multimodal Visual Dictionary Learning

2013-06-01 · CVPR 2013 6 · Go Irie, Dong Liu, Zhenguo Li, Shih-Fu Chang

nary learning methods rely on image descriptors alone or together with class labels. However, Web images are often associated with text data which may carry substantial information regarding image semantics, and may be e…

Bayesian InferenceClusteringDictionary LearningImage Categorization+1