Augmenting Vision-Based Human Pose Estimation with Rotation Matrix
Fitness applications are commonly used to monitor activities within the gym, but they often fail to automatically track indoor activities inside the gym. This study proposes a model that utilizes pose estimation combined with a novel data augmentation method, i.e., rotation matrix. We aim to enhance the classification accuracy of activity recognition based on pose estimation data. Through our experiments, we experiment with different classification algorithms along with image augmentation approaches. Our findings demonstrate that the SVM with SGD optimization, using data augmentation with the Rotation Matrix, yields the most accurate results, achieving a 96% accuracy rate in classifying five physical activities. Conversely, without implementing the data augmentation techniques, the baseline accuracy remains at a modest 64%.
Code (0)
등록된 구현이 없습니다.
Tasks
Activity RecognitionData AugmentationImage AugmentationPose EstimationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
HoloFace: Augmenting Human-to-Human Interactions on HoloLens
We present HoloFace, an open-source framework for face alignment, head pose estimation and facial attribute retrieval for Microsoft HoloLens. HoloFace implements two state-of-the-art face alignment methods which can be u…
AttributeEmotion RecognitionFace AlignmentFace Model+3Learning Unorthogonalized Matrices for Rotation Estimation
Estimating 3D rotations is a common procedure for 3D computer vision. The accuracy depends heavily on the rotation representation. One form of representation -- rotation matrices -- is popular due to its continuity, espe…
3D Human Pose EstimationPose EstimationJoint Hand Detection and Rotation Estimation by Using CNN
Hand detection is essential for many hand related tasks, e.g. parsing hand pose, understanding gesture, which are extremely useful for robotics and human-computer interaction. However, hand detection in uncontrolled envi…
General ClassificationHand Detectionobject-detectionObject DetectionOn the Role of Rotation Equivariance in Monocular 2D-to-3D Human Pose Lifting
Estimating 3D from 2D is one of the central tasks in computer vision. In this work, we consider the monocular setting, i.e. single-view input, for 3D human pose estimation (HPE), where the goal is to predict a 3D point s…
3D Human Pose EstimationKeypoint DetectionData AugmentationSPIN: Simplifying Polar Invariance for Neural networks Application to vision-based irradiance forecasting
Translational invariance induced by pooling operations is an inherent property of convolutional neural networks, which facilitates numerous computer vision tasks such as classification. Yet to leverage rotational invaria…
Data AugmentationSolar Irradiance Forecasting