paper-with-me

홈 › Papers

Effective Actor-centric Human-object Interaction Detection

2022-02-24 · Kunlun Xu, Zhimin Li, Zhijun Zhang, Leizhen Dong, Wenhui Xu, Luxin Yan, Sheng Zhong, Xu Zou

While Human-Object Interaction(HOI) Detection has achieved tremendous advances in recent, it still remains challenging due to complex interactions with multiple humans and objects occurring in images, which would inevitably lead to ambiguities. Most existing methods either generate all human-object pair candidates and infer their relationships by cropped local features successively in a two-stage manner, or directly predict interaction points in a one-stage procedure. However, the lack of spatial configurations or reasoning steps of two- or one- stage methods respectively limits their performance in such complex scenes. To avoid this ambiguity, we propose a novel actor-centric framework. The main ideas are that when inferring interactions: 1) the non-local features of the entire image guided by actor position are obtained to model the relationship between the actor and context, and then 2) we use an object branch to generate pixel-wise interaction area prediction, where the interaction area denotes the object central area. Moreover, we also use an actor branch to get interaction prediction of the actor and propose a novel composition strategy based on center-point indexing to generate the final HOI prediction. Thanks to the usage of the non-local features and the partly-coupled property of the human-objects composition strategy, our proposed framework can detect HOI more accurately especially for complex images. Extensive experimental results show that our method achieves the state-of-the-art on the challenging V-COCO and HICO-DET benchmarks and is more robust especially in multiple persons and/or objects scenes.

📄 PDF Abstract BibTeX arXiv:2202.11998

Code (0)

등록된 구현이 없습니다.

Tasks

Human-Object Interaction DetectionObject

Similar Papers 제목 키워드 기반

Object-centric Video Representation for Long-term Action Anticipation

2023-10-31 · Ce Zhang, Changcheng Fu, Shijie Wang, Nakul Agarwal 외

This paper focuses on building object-centric representations for long-term action anticipation in videos. Our key motivation is that objects provide important cues to recognize and predict human-object interactions, esp…

Action AnticipationHuman-Object Interaction DetectionLong Term Action AnticipationObject+2

EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views

2024-05-22 · Yuhang Yang, Wei Zhai, Chengfeng Wang, Chengjun Yu 외

Understanding egocentric human-object interaction (HOI) is a fundamental aspect of human-centric perception, facilitating applications like AR/VR and embodied AI. For the egocentric HOI, in addition to perceiving semanti…

Human-Object Interaction DetectionObject

Deep Dual Relation Modeling for Egocentric Interaction Recognition

2019-05-31 · CVPR 2019 6 · Haoxin Li, Yijun Cai, Wei-Shi Zheng

Egocentric interaction recognition aims to recognize the camera wearer's interactions with the interactor who faces the camera wearer in egocentric videos. In such a human-human interaction analysis problem, it is crucia…

Relation

Mask2IV: Interaction-Centric Video Generation via Mask Trajectories

2025-10-03 · Gen Li, Bo Zhao, Jianfei Yang, Laura Sevilla-Lara arxiv

Generating interaction-centric videos, such as those depicting humans or robots interacting with objects, is crucial for embodied intelligence, as they provide rich and diverse visual priors for robot learning, manipulat…

Video Generation

The MECCANO Dataset: Understanding Human-Object Interactions from Egocentric Videos in an Industrial-like Domain

2020-10-12 · Francesco Ragusa, Antonino Furnari, Salvatore Livatino, Giovanni Maria Farinella

Wearable cameras allow to collect images and videos of humans interacting with the world. While human-object interactions have been thoroughly investigated in third person vision, the problem has been understudied in ego…

Action RecognitionActive Object DetectionHuman-Object Interaction DetectionObject+3