From Detection to Action Recognition: An Edge-Based Pipeline for Robot Human Perception
Mobile service robots are proving to be increasingly effective in a range of applications, such as healthcare, monitoring Activities of Daily Living (ADL), and facilitating Ambient Assisted Living (AAL). These robots heavily rely on Human Action Recognition (HAR) to interpret human actions and intentions. However, for HAR to function effectively on service robots, it requires prior knowledge of human presence (human detection) and identification of individuals to monitor (human tracking). In this work, we propose an end-to-end pipeline that encompasses the entire process, starting from human detection and tracking, leading to action recognition. The pipeline is designed to operate in near real-time while ensuring all stages of processing are performed on the edge, reducing the need for centralised computation. To identify the most suitable models for our mobile robot, we conducted a series of experiments comparing state-of-the-art solutions based on both their detection performance and efficiency. To evaluate the effectiveness of our proposed pipeline, we proposed a dataset comprising daily household activities. By presenting our findings and analysing the results, we demonstrate the efficacy of our approach in enabling mobile robots to understand and respond to human behaviour in real-world scenarios relying mainly on the data from their RGB cameras.
Code (0)
등록된 구현이 없습니다.
Tasks
Action RecognitionHuman DetectionTemporal Action LocalizationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Exploring Emerging Trends and Research Opportunities in Visual Place Recognition
Visual-based recognition, e.g., image classification, object detection, etc., is a long-standing challenge in computer vision and robotics communities. Concerning the roboticists, since the knowledge of the environment i…
image-classificationImage ClassificationLoop Closure Detectionobject-detection+3FRAME: Fast and Robust Autonomous 3D point cloud Map-merging for Egocentric multi-robot exploration
This article presents a 3D point cloud map-merging framework for egocentric heterogeneous multi-robot exploration, based on overlap detection and alignment, that is independent of a manual initial guess or prior knowledg…
Point Cloud RegistrationIndoor Semantic Scene Understanding using Multi-modality Fusion
Seamless Human-Robot Interaction is the ultimate goal of developing service robotic systems. For this, the robotic agents have to understand their surroundings to better complete a given task. Semantic scene understandin…
Scene UnderstandingReal-Time Object Detection and Recognition on Low-Compute Humanoid Robots using Deep Learning
We envision that in the near future, humanoid robots would share home space and assist us in our daily and routine activities through object manipulations. One of the fundamental technologies that need to be developed fo…
object-detectionObject DetectionQuantizationReal-Time Object DetectionReal-Time Multimodal Signal Processing for HRI in RoboCup: Understanding a Human Referee
Advancing human-robot communication is crucial for autonomous systems operating in dynamic environments, where accurate real-time interpretation of human signals is essential. RoboCup provides a compelling scenario for t…
Gesture Recognition