paper-with-me

홈 › Papers

Learning secondary tool affordances of human partners using iCub robot's egocentric data

2024-07-16 · Bosong Ding, Erhan Oztop, Giacomo Spigler, Murat Kirtay

Objects, in particular tools, provide several action possibilities to the agents that can act on them, which are generally associated with the term of affordances. A tool is typically designed for a specific purpose, such as driving a nail in the case of a hammer, which we call as the primary affordance. A tool can also be used beyond its primary purpose, in which case we can associate this auxiliary use with the term secondary affordance. Previous work on affordance perception and learning has been mostly focused on primary affordances. Here, we address the less explored problem of learning the secondary tool affordances of human partners. To do this, we use the iCub robot to observe human partners with three cameras while they perform actions on twenty objects using four different tools. In our experiments, human partners utilize tools to perform actions that do not correspond to their primary affordances. For example, the iCub robot observes a human partner using a ruler for pushing, pulling, and moving objects instead of measuring their lengths. In this setting, we constructed a dataset by taking images of objects before and after each action is executed. We then model learning secondary affordances by training three neural networks (ResNet-18, ResNet-50, and ResNet-101) each on three tasks, using raw images showing the initial' and final' position of objects as input: (1) predicting the tool used to move an object, (2) predicting the tool used with an additional categorical input that encoded the action performed, and (3) joint prediction of both tool used and action performed. Our results indicate that deep learning architectures enable the iCub robot to predict secondary tool affordances, thereby paving the road for human-robot collaborative object manipulation involving complex affordances.

📄 PDF Abstract BibTeX arXiv:2407.11922

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

iCub! Do you recognize what I am doing?: multimodal human action recognition on multisensory-enabled iCub robot

2022-12-17 · Kas Kniesmeijer, Murat Kirtay

This study uses multisensory data (i.e., color and depth) to recognize human actions in the context of multimodal human-robot interaction. Here we employed the iCub robot to observe the predefined actions of the human pa…

Action RecognitionEnsemble LearningTemporal Action Localization

Learning at the Ends: From Hand to Tool Affordances in Humanoid Robots

2018-04-09 · Giovanni Saponaro, Pedro Vicente, Atabak Dehban, Lorenzo Jamone 외

One of the open challenges in designing robots that operate successfully in the unpredictable human environment is how to make them able to predict what actions they can perform on objects, and what their effects will be…

Decision Making

Privacy in Human-AI Romantic Relationships: Concerns, Boundaries, and Agency

2026-01-23 · Rongjun Ma, Shijing He, Jose Luis Martin-Navarro, Xiao Zhan 외 arxiv

An increasing number of LLM-based applications are being developed to facilitate romantic relationships with AI partners, yet the safety and privacy risks in these partnerships remain largely underexplored. In this work,…

From Particles to Agents: Hallucination as a Metric for Cognitive Friction in Spatial Simulation

2026-01-29 · Javier Argota Sánchez-Vaquerizo, Luis Borunda Monsivais arxiv

Traditional architectural simulations (e.g. Computational Fluid Dynamics, evacuation, structural analysis) model elements as deterministic physics-based "particles" rather than cognitive "agents". To bridge this, we intr…

Spatial Reasoning

iCub World: Friendly Robots Help Building Good Vision Data-Sets

2013-06-15 · Sean Ryan Fanello, Carlo Ciliberto, Matteo Santoro, Lorenzo Natale 외

In this paper we present and start analyzing the iCub World data-set, an object recognition data-set, we acquired using a Human-Robot Interaction (HRI) scheme and the iCub humanoid robot platform. Our set up allows for r…

Object Recognition