Object Activity Scene Description, Construction and Recognition
Action recognition is a critical task for social robots to meaningfully engage with their environment. 3D human skeleton-based action recognition is an attractive research area in recent years. Although, the existing approaches are good at action recognition, it is a great challenge to recognize a group of actions in an activity scene. To tackle this problem, at first, we partition the scene into several primitive actions (PAs) based upon motion attention mechanism. Then, the primitive actions are described by the trajectory vectors of corresponding joints. After that, motivated by text classification based on word embedding, we employ convolution neural network (CNN) to recognize activity scenes by considering motion of joints as "word" of activity. The experimental results on the scenes of human activity dataset show the efficiency of the proposed approach.
Code (0)
등록된 구현이 없습니다.
Tasks
Action RecognitionGeneral ClassificationObjectSkeleton Based Action RecognitionTemporal Action Localizationtext-classificationText ClassificationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Prediction and Description of Near-Future Activities in Video
Most of the existing works on human activity analysis focus on recognition or early recognition of the activity labels from complete or partial observations. Similarly, almost all of the existing video captioning approac…
PredictionVideo CaptioningVideo DescriptionBonn Activity Maps: Dataset Description
The key prerequisite for accessing the huge potential of current machine learning techniques is the availability of large databases that capture the complex relations of interest. Previous datasets are focused on either …
Activity RecognitionNeurons: Emulating the Human Visual Cortex Improves Fidelity and Interpretability in fMRI-to-Video Reconstruction
Decoding visual stimuli from neural activity is essential for understanding the human brain. While fMRI methods have successfully reconstructed static images, fMRI-to-video reconstruction faces challenges due to the need…
Semantic SegmentationVideo ReconstructionTopological map construction and scene recognition for vehicle localization
This paper presents a vehicle localization method to assist vehicle navigation based on topological map construction and scene recognition. A topological map is constructed using omni-directional image sequences, and …
Change DetectionImage RetrievalRetrievalScene Change Detection+1Efficient data-driven encoding of scene motion using Eccentricity
This paper presents a novel approach of representing dynamic visual scenes with static maps generated from video/image streams. Such representation allows easy visual assessment of motion in dynamic environments. These m…
Activity RecognitionIntent RecognitionObject TrackingVideo Description