paper-with-me

홈 › Papers

Proposition of Affordance-Driven Environment Recognition Framework Using Symbol Networks in Large Language Models

2025-04-02 · Kazuma Arii, Satoshi Kurihara

In the quest to enable robots to coexist with humans, understanding dynamic situations and selecting appropriate actions based on common sense and affordances are essential. Conventional AI systems face challenges in applying affordance, as it represents implicit knowledge derived from common sense. However, large language models (LLMs) offer new opportunities due to their ability to process extensive human knowledge. This study proposes a method for automatic affordance acquisition by leveraging LLM outputs. The process involves generating text using LLMs, reconstructing the output into a symbol network using morphological and dependency analysis, and calculating affordances based on network distances. Experiments using ``apple'' as an example demonstrated the method's ability to extract context-dependent affordances with high explainability. The results suggest that the proposed symbol network, reconstructed from LLM outputs, enables robots to interpret affordances effectively, bridging the gap between symbolized data and human-like situational understanding.

📄 PDF Abstract BibTeX arXiv:2504.01644

Code (0)

등록된 구현이 없습니다.

Tasks

Common Sense Reasoning

Similar Papers 제목 키워드 기반

Text-driven object affordance for guiding grasp-type recognition in multimodal robot teaching

2021-02-27 · Naoki Wake, Daichi Saito, Kazuhiro Sasabuchi, Hideki Koike 외

This study investigates how text-driven object affordance, which provides prior knowledge about grasp types for each object, affects image-based grasp-type recognition in robot teaching. The researchers created labeled d…

Mixed RealityObjectVocal Bursts Type Prediction

Grasp-type Recognition Leveraging Object Affordance

2020-08-26 · Wake Naoki, Sasabuchi Kazuhiro, Ikeuchi Katsushi

A key challenge in robot teaching is grasp-type recognition with a single RGB image and a target object name. Here, we propose a simple yet effective pipeline to enhance learning-based recognition by leveraging a prior d…

ObjectVocal Bursts Type Prediction

Grounding 3D Scene Affordance From Egocentric Interactions

2024-09-29 · Cuiyu Liu, Wei Zhai, Yuhang Yang, Hongchen Luo 외

Grounding 3D scene affordance aims to locate interactive regions in 3D environments, which is crucial for embodied agents to interact intelligently with their surroundings. Most existing approaches achieve this by mappin…

ReSemAct: Advancing Fine-Grained Robotic Manipulation via Semantic Structuring and Affordance Refinement

2025-07-24 · Chenyu Su, Weiwei Shang, Chen Qian, Fei Zhang 외 arxiv

Fine-grained robotic manipulation requires grounding natural language into appropriate affordance targets. However, most existing methods driven by foundation models often compress rich semantics into oversimplified affo…

An Interactive Navigation Method with Effect-oriented Affordance

2024-01-01 · CVPR 2024 1 · Xiaohan Wang, Yuehu Liu, Xinhang Song, Yuyi Liu 외

Visual navigation is to let the agent reach the target according to the continuous visual input. In most previous works visual navigation is usually assumed to be done in a static and ideal environment: the target is…

NavigateVisual Navigation