Video-to-image Affordance Grounding
3개 벤치마크 · 논문 4편 · 이 태스크의 논문 보기 →
Benchmarks
Most implemented
Affordance Grounding from Demonstration Video to Target Image
Grounded Human-Object Interaction Hotspots from Video
Papers
Affordance Grounding from Demonstration Video to Target Image
Humans excel at learning from expert demonstrations and solving their own problems. To equip intelligent robots and assistants, such as AR glasses, with this ability, it is essential to ground human hand interactions (i.…
DecoderVideo-to-image Affordance GroundingLearning Visual Affordance Grounding from Demonstration Videos
Visual affordance grounding aims to segment all possible interaction regions between people and objects from an image/video, which is beneficial for many applications, such as robot grasping and action recognition. Howev…
Action RecognitionObjectVideo-to-image Affordance GroundingGrounded Human-Object Interaction Hotspots from Video
Learning how to interact with objects is an important step towards embodied visual intelligence, but existing techniques suffer from heavy supervision or sensing requirements. We propose an approach to learn human-object…
Human-Object Interaction DetectionObjectObject RecognitionSemantic Segmentation+1Demo2Vec: Reasoning Object Affordances From Online Videos
Watching expert demonstrations is an important way for humans and robots to reason about affordances of unseen objects. In this paper, we consider the problem of reasoning object affordances through the feature embedding…
ObjectVideo-to-image Affordance Grounding