paper-with-me

Papers

Context-driven Visual Object Recognition based on Knowledge Graphs

2022-10-20 · Sebastian Monka, Lavdim Halilaj, Achim Rettinger

Current deep learning methods for object recognition are purely data-driven and require a large number of training samples to achieve good results. Due to their sole dependence on image data, these methods tend to fail when confronted with new environments where even small deviations occur. Human perception, however, has proven to be significantly more robust to such distribution shifts. It is assumed that their ability to deal with unknown scenarios is based on extensive incorporation of contextual knowledge. Context can be based either on object co-occurrences in a scene or on memory of experience. In accordance with the human visual cortex which uses context to form different object representations for a seen image, we propose an approach that enhances deep learning methods by using external contextual knowledge encoded in a knowledge graph. Therefore, we extract different contextual views from a generic knowledge graph, transform the views into vector space and infuse it into a DNN. We conduct a series of experiments to investigate the impact of different contextual views on the learned object representations for the same image dataset. The experimental results provide evidence that the contextual views influence the image representations in the DNN differently and therefore lead to different predictions for the same images. We also show that context helps to strengthen the robustness of object recognition models for out-of-distribution images, usually occurring in transfer learning tasks or real-world scenarios.

📄 PDF Abstract BibTeX arXiv:2210.11233

Code (0)

등록된 구현이 없습니다.

Tasks

Knowledge GraphsObjectObject RecognitionTransfer Learning

Similar Papers 제목 키워드 기반

The functional role of cue-driven feature-based feedback in object recognition

2019-03-25 · Sushrut Thorat, Marcel van Gerven, Marius Peelen

Visual object recognition is not a trivial task, especially when the objects are degraded or surrounded by clutter or presented briefly. External cues (such as verbal cues or visual context) can boost recognition perform…

Object Recognition

SkeletonContext: Skeleton-side Context Prompt Learning for Zero-Shot Skeleton-based Action Recognition

2026-03-31 · Ning Wang, Tieyue Wu, Naeha Sharif, Farid Boussaid 외 arxiv

Zero-shot skeleton-based action recognition aims to recognize unseen actions by transferring knowledge from seen categories through semantic descriptions. Most existing methods typically align skeleton features with text…

Action UnderstandingAction Recognition

Knowledge-driven Scene Priors for Semantic Audio-Visual Embodied Navigation

2022-12-21 · Gyan Tatiya, Jonathan Francis, Luca Bondi, Ingrid Navarro 외

Generalisation to unseen contexts remains a challenge for embodied navigation agents. In the context of semantic audio-visual navigation (SAVi) tasks, the notion of generalisation should include both generalising to unse…

Visual Navigation

Text-driven object affordance for guiding grasp-type recognition in multimodal robot teaching

2021-02-27 · Naoki Wake, Daichi Saito, Kazuhiro Sasabuchi, Hideki Koike 외

This study investigates how text-driven object affordance, which provides prior knowledge about grasp types for each object, affects image-based grasp-type recognition in robot teaching. The researchers created labeled d…

Mixed RealityObjectVocal Bursts Type Prediction

PVLR: Prompt-driven Visual-Linguistic Representation Learning for Multi-Label Image Recognition

2024-01-31 · Hao Tan, Zichang Tan, Jun Li, Jun Wan 외

Multi-label image recognition is a fundamental task in computer vision. Recently, vision-language models have made notable advancements in this area. However, previous methods often failed to effectively leverage the ric…

Multi-Label Image RecognitionRepresentation Learning