Affordance Recognition
2개 벤치마크 · 논문 14편 · 이 태스크의 논문 보기 →
Benchmarks
HICO-DET
Most implemented
Visual Compositional Learning for Human-Object Interaction Detection
Affordance Transfer Learning for Human-Object Interaction Detection
3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians
Detecting Human-Object Interaction via Fabricated Compositional Learning
Papers
ENACT: Evaluating Embodied Cognition with World Modeling of Egocentric Interaction
Embodied cognition argues that intelligence arises from sensorimotor interaction rather than passive observation. It raises an intriguing question: do modern vision-language models (VLMs), trained largely in a disembodie…
Visual Question AnsweringAffordance RecognitionRoboAfford++: A Generative AI-Enhanced Dataset for Multimodal Affordance Learning in Robotic Manipulation and Navigation
Robotic manipulation and navigation are fundamental capabilities of embodied intelligence, enabling effective robot interactions with the physical world. Achieving these capabilities requires a cohesive understanding of …
Affordance RecognitionScene UnderstandingObject RecognitionQuestion AnsweringObject Affordance Recognition and Grounding via Multi-scale Cross-modal Representation Learning
A core problem of Embodied AI is to learn object manipulation from observation, as humans do. To achieve this, it is important to localize 3D object affordance areas through observation such as images (3D affordance grou…
Representation LearningAffordance Recognition3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians
3D affordance reasoning is essential in associating human instructions with the functional regions of 3D objects, facilitating precise, task-oriented manipulations in embodied AI. However, current methods, which predomin…
3DGSAffordance RecognitionFine-grained Affordance Annotation for Egocentric Hand-Object Interaction Videos
Object affordance is an important concept in hand-object interaction, providing information on action possibilities based on human motor capacity and objects' physical property thus benefiting tasks such as action antici…
Action AnticipationAction RecognitionAffordance RecognitionHuman-Object Interaction Detection+2AROS: Affordance Recognition with One-Shot Human Stances
We present AROS, a one-shot learning approach that uses an explicit representation of interactions between highly-articulated human poses and 3D scenes. The approach is one-shot as the method does not require re-training…
Affordance RecognitionOne-Shot Learning