paper-with-me

Papers

Visual Affordance and Function Understanding: A Survey

2018-07-18 · Mohammed Hassanin, Salman Khan, Murat Tahtali

Nowadays, robots are dominating the manufacturing, entertainment and healthcare industries. Robot vision aims to equip robots with the ability to discover information, understand it and interact with the environment. These capabilities require an agent to effectively understand object affordances and functionalities in complex visual domains. In this literature survey, we first focus on Visual affordances and summarize the state of the art as well as open problems and research gaps. Specifically, we discuss sub-problems such as affordance detection, categorization, segmentation and high-level reasoning. Furthermore, we cover functional scene understanding and the prevalent functional descriptors used in the literature. The survey also provides necessary background to the problem, sheds light on its significance and highlights the existing challenges for affordance and functionality learning.

📄 PDF Abstract BibTeX arXiv:1807.06775

Code (0)

등록된 구현이 없습니다.

Tasks

Affordance DetectionScene UnderstandingSurvey

Similar Papers 제목 키워드 기반

3D AffordanceNet: A Benchmark for Visual Object Affordance Understanding

2021-03-30 · CVPR 2021 1 · Shengheng Deng, Xun Xu, Chaozheng Wu, Ke Chen 외

The ability to understand the ways to interact with objects from visual cues, a.k.a. visual affordance, is essential to vision-guided robotic research. This involves categorizing, segmenting and reasoning of visual affor…

Affordance DetectionBenchmarkingObject

Mining Semantic Affordances of Visual Object Categories

2015-06-01 · CVPR 2015 6 · Yu-Wei Chao, Zhan Wang, Rada Mihalcea, Jia Deng

Affordances are fundamental attributes of objects. Affordances reveal the functionalities of objects and the possible actions that can be performed on them. Understanding affordances is crucial for recognizing human acti…

Collaborative FilteringObject

Indoor Scene Understanding in 2.5/3D for Autonomous Agents: A Survey

2018-03-09 · Muzammal Naseer, Salman H. Khan, Fatih Porikli

With the availability of low-cost and compact 2.5/3D visual sensing devices, computer vision community is experiencing a growing interest in visual scene understanding of indoor environments. This survey paper provides a…

3D Reconstructionobject-detectionObject DetectionPose Estimation+5

AFUN: Towards an Affordance Foundation Model for Functionality Understanding

2026-06-01 · Zhaoning Wang, Yi Zhong, Jiawei Fu, Henrik I. Christensen 외 arxiv

Affordance understanding bridges visual perception and physical action, serving as an explainable interface for robot manipulation in open and unstructured real-world environments. Yet, building an affordance foundation …

Robot Manipulation

What Objects Enable, Not What They Are: Functional Latent Spaces for Affordance Reasoning

2026-06-04 · Rohan Siva, Neel P. Bhatt, Yunhao Yang, Seoyoung Lee 외 arxiv

Existing robot planning systems rely on appearance-based reasoning, where visual observations are encoded into latent spaces organized around object appearances (e.g., recognizing a "cart" based on how it looks). However…