Human Interaction Recognition Framework based on Interacting Body Part Attention
Human activity recognition in videos has been widely studied and has recently gained significant advances with deep learning approaches; however, it remains a challenging task. In this paper, we propose a novel framework that simultaneously considers both implicit and explicit representations of human interactions by fusing information of local image where the interaction actively occurred, primitive motion with the posture of individual subject's body parts, and the co-occurrence of overall appearance change. Human interactions change, depending on how the body parts of each human interact with the other. The proposed method captures the subtle difference between different interactions using interacting body part attention. Semantically important body parts that interact with other objects are given more weight during feature representation. The combined feature of interacting body part attention-based individual representation and the co-occurrence descriptor of the full-body appearance change is fed into long short-term memory to model the temporal dynamics over time in a single framework. We validate the effectiveness of the proposed method using four widely used public datasets by outperforming the competing state-of-the-art method.
Code (0)
등록된 구현이 없습니다.
Tasks
Activity RecognitionActivity Recognition In VideosHuman Activity RecognitionHuman Interaction RecognitionSimilar Papers 제목 키워드 기반
Transformer-based Action recognition in hand-object interacting scenarios
This report describes the 2nd place solution to the ECCV 2022 Human Body, Hands, and Activities (HBHA) from Egocentric and Multi-view Cameras Challenge: Action Recognition. This challenge aims to recognize hand-object in…
Action RecognitionObjectRobot self/other distinction: active inference meets neural networks learning in a mirror
Self/other distinction and self-recognition are important skills for interacting with the world, as it allows humans to differentiate own actions from others and be self-aware. However, only a selected group of animals, …
Learning Human-Object Interaction for 3D Human Pose Estimation from LiDAR Point Clouds
Understanding humans from LiDAR point clouds is one of the most critical tasks in autonomous driving due to its close relationships with pedestrian safety, yet it remains challenging in the presence of diverse human-obje…
3D Human Pose EstimationContrastive LearningAutonomous DrivingPoint CloudsTowards Social Artificial Intelligence: Nonverbal Social Signal Prediction in A Triadic Interaction
We present a new research task and a dataset to understand human social interactions via computational methods, to ultimately endow machines with the ability to encode and decode a broad channel of social signals humans …
InterCap: Joint Markerless 3D Tracking of Humans and Objects in Interaction
Humans constantly interact with daily objects to accomplish tasks. To understand such interactions, computers need to reconstruct these from cameras observing whole-body interaction with scenes. This is challenging due t…
ObjectPose Estimation