paper-with-me

Papers

Semantic Linking Maps for Active Visual Object Search

2020-06-18 · Zhen Zeng, Adrian Röfer, Odest Chadwicke Jenkins

We aim for mobile robots to function in a variety of common human environments. Such robots need to be able to reason about the locations of previously unseen target objects. Landmark objects can help this reasoning by narrowing down the search space significantly. More specifically, we can exploit background knowledge about common spatial relations between landmark and target objects. For example, seeing a table and knowing that cups can often be found on tables aids the discovery of a cup. Such correlations can be expressed as distributions over possible pairing relationships of objects. In this paper, we propose an active visual object search strategy method through our introduction of the Semantic Linking Maps (SLiM) model. SLiM simultaneously maintains the belief over a target object's location as well as landmark objects' locations, while accounting for probabilistic inter-object spatial relations. Based on SLiM, we describe a hybrid search strategy that selects the next best view pose for searching for the target object based on the maintained belief. We demonstrate the efficiency of our SLiM-based search strategy through comparative experiments in simulated environments. We further demonstrate the real-world applicability of SLiM-based search in scenarios with a Fetch mobile manipulation robot.

📄 PDF Abstract BibTeX arXiv:2006.10807

Code (0)

등록된 구현이 없습니다.

Tasks

Object

Similar Papers 제목 키워드 기반

Linking in Style: Understanding learned features in deep learning models

2024-09-25 · Maren H. Wehrheim, Pamela Osuna-Vargas, Matthias Kaschube

Convolutional neural networks (CNNs) learn abstract features to perform object classification, but understanding these features remains challenging due to difficult-to-interpret results or high computational costs. We pr…

counterfactualDeep Learning

Reverse Region-to-Entity Annotation for Pixel-Level Visual Entity Linking

2024-12-18 · Zhengfei Xu, Sijia Zhao, Yanchao Hao, Xiaolong Liu 외

Visual Entity Linking (VEL) is a crucial task for achieving fine-grained visual understanding, matching objects within images (visual mentions) to entities in a knowledge base. Previous VEL tasks rely on textual inputs, …

Entity Linking

A Graph-based Interactive Reasoning for Human-Object Interaction Detection

2020-07-14 · Dongming Yang, YueXian Zou

Human-Object Interaction (HOI) detection devotes to learn how humans interact with surrounding objects via inferring triplets of < human, verb, object >. However, recent HOI detection methods mostly rely on additional an…

Human-Object Interaction Detection

RareSpot+: A Benchmark, Model, and Active Learning Framework for Small and Rare Wildlife in Aerial Imagery

2026-04-21 · Bowen Zhang, Jesse T. Boulerice, Charvi Mendiratta, Nikhil Kuniyil 외 arxiv

Automated wildlife monitoring from aerial imagery is vital for conservation but remains limited by two persistent challenges: the difficulty of detecting small, rare species and the high cost of large-scale expert annota…

Active Learning

DWE+: Dual-Way Matching Enhanced Framework for Multimodal Entity Linking

2024-04-07 · Shezheng Song, Shasha Li, Shan Zhao, Xiaopeng Li 외

Multimodal entity linking (MEL) aims to utilize multimodal information (usually textual and visual information) to link ambiguous mentions to unambiguous entities in knowledge base. Current methods facing main issues: (1…

Contrastive LearningEntity Linking