paper-with-me

홈 › Papers

Phrase-Based Affordance Detection via Cyclic Bilateral Interaction

2022-02-24 · Liangsheng Lu, Wei Zhai, Hongchen Luo, Yu Kang, Yang Cao

Affordance detection, which refers to perceiving objects with potential action possibilities in images, is a challenging task since the possible affordance depends on the person's purpose in real-world application scenarios. The existing works mainly extract the inherent human-object dependencies from image/video to accommodate affordance properties that change dynamically. In this paper, we explore to perceive affordance from a vision-language perspective and consider the challenging phrase-based affordance detection problem,i.e., given a set of phrases describing the action purposes, all the object regions in a scene with the same affordance should be detected. To this end, we propose a cyclic bilateral consistency enhancement network (CBCE-Net) to align language and vision features progressively. Specifically, the presented CBCE-Net consists of a mutual guided vision-language module that updates the common features of vision and language in a progressive manner, and a cyclic interaction module (CIM) that facilitates the perception of possible interaction with objects in a cyclic manner. In addition, we extend the public Purpose-driven Affordance Dataset (PAD) by annotating affordance categories with short phrases. The contrastive experimental results demonstrate the superiority of our method over nine typical methods from four relevant fields in terms of both objective metrics and visual quality. The related code and dataset will be released at \url{https://github.com/lulsheng/CBCE-Net}.

📄 PDF Abstract BibTeX arXiv:2202.12076

Code (4)

lhc1224/OSAD_Net 공식 구현 pytorch
lulsheng/cbce-net 공식 구현 tf
lhc1224/cross-view-affordance-grounding pytorch
lhc1224/cross-view-ag pytorch

Tasks

Affordance Detection

Similar Papers 제목 키워드 기반

ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection

2026-05-06 · Minh Anh Nguyen, Quang Huy Tran, Bao Ngoc Le, SuiYang Guang 외 arxiv

Open-vocabulary human-object interaction (HOI) detection requires recognizing interaction phrases that may not appear as annotated categories during training. Recent vision-language HOI detectors improve semantic transfe…

Human-Object Interaction Detection

Affordance Transfer Learning for Human-Object Interaction Detection

2021-04-07 · CVPR 2021 1 · Zhi Hou, Baosheng Yu, Yu Qiao, Xiaojiang Peng 외

Reasoning the human-object interactions (HOI) is essential for deeper scene understanding, while object affordances (or functionalities) are of great importance for human to discover unseen HOIs with novel objects. Inspi…

Affordance DetectionAffordance RecognitionHuman-Object Interaction Concept DiscoveryHuman-Object Interaction Detection+3

Multi-label affordance mapping from egocentric vision

2023-09-05 · ICCV 2023 1 · Lorenzo Mur-Labadia, Jose J. Guerrero, Ruben Martinez-Cantin

Accurate affordance detection and segmentation with pixel precision is an important piece in many complex systems based on interactions, such as robots and assitive devices. We present a new approach to affordance percep…

Affordance DetectionSegmentation

3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds

2025-02-27 · Hengshuo Chu, Xiang Deng, Qi Lv, Xiaoyang Chen 외

3D Affordance detection is a challenging problem with broad applications on various robotic tasks. Existing methods typically formulate the detection paradigm as a label-based semantic segmentation task. This paradigm re…

Affordance DetectionHuman-Object Interaction DetectionSegmentationSemantic Segmentation+1

Egocentric affordance detection with the one-shot geometry-driven Interaction Tensor

2019-06-13 · Eduardo Ruiz, Walterio Mayol-Cuevas

In this abstract we describe recent [4,7] and latest work on the determination of affordances in visually perceived 3D scenes. Our method builds on the hypothesis that geometry on its own provides enough information to e…

Affordance Detection