paper-with-me

홈 › Papers

Commonsense Scene Semantics for Cognitive Robotics: Towards Grounding Embodied Visuo-Locomotive Interactions

2017-09-15 · Jakob Suchan, Mehul Bhatt

We present a commonsense, qualitative model for the semantic grounding of embodied visuo-spatial and locomotive interactions. The key contribution is an integrative methodology combining low-level visual processing with high-level, human-centred representations of space and motion rooted in artificial intelligence. We demonstrate practical applicability with examples involving object interactions, and indoor movement.

📄 PDF Abstract BibTeX arXiv:1709.05293

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Combining Commonsense Reasoning and Knowledge Acquisition to Guide Deep Learning in Robotics

2022-01-25 · Mohan Sridharan, Tiago Mota

Algorithms based on deep network models are being used for many pattern recognition and decision-making tasks in robotics and AI. Training these models requires a large labeled dataset and considerable computational reso…

Decision MakingLogical Reasoning

Grounding Dynamic Spatial Relations for Embodied (Robot) Interaction

2016-07-26 · Michael Spranger, Jakob Suchan, Mehul Bhatt, Manfred Eppe

This paper presents a computational model of the processing of dynamic spatial relations occurring in an embodied robotic interaction setup. A complete system is introduced that allows autonomous robots to produce and in…

Deep Semantic Abstractions of Everyday Human Activities: On Commonsense Representations of Human Interactions

2017-10-10 · Jakob Suchan, Mehul Bhatt

We propose a deep semantic characterization of space and motion categorically from the viewpoint of grounding embodied human-object interactions. Our key focus is on an ontological model that would be adept to formalisat…

Human-Object Interaction DetectionObjectRelational Reasoning

LLM-Grounder: Open-Vocabulary 3D Visual Grounding with Large Language Model as an Agent

2023-09-21 · Jianing Yang, Xuweiyi Chen, Shengyi Qian, Nikhil Madaan 외

3D visual grounding is a critical skill for household robots, enabling them to navigate, manipulate objects, and answer questions based on their environment. While existing approaches often rely on extensive labeled data…

3D visual groundingLanguage ModelingLanguage ModellingLarge Language Model+3

DenseGrounding: Improving Dense Language-Vision Semantics for Ego-Centric 3D Visual Grounding

2025-05-08 · Henry Zheng, Hao Shi, Qihang Peng, Yong Xien Chng 외

Enabling intelligent agents to comprehend and interact with 3D environments through natural language is crucial for advancing robotics and human-computer interaction. A fundamental task in this field is ego-centric 3D vi…

3D visual groundingcross-modal alignmentVisual Grounding