Sparse and Privacy-enhanced Representation for Human Pose Estimation
We propose a sparse and privacy-enhanced representation for Human Pose Estimation (HPE). Given a perspective camera, we use a proprietary motion vector sensor(MVS) to extract an edge image and a two-directional motion vector image at each time frame. Both edge and motion vector images are sparse and contain much less information (i.e., enhancing human privacy). We advocate that edge information is essential for HPE, and motion vectors complement edge information during fast movements. We propose a fusion network leveraging recent advances in sparse convolution used typically for 3D voxels to efficiently process our proposed sparse representation, which achieves about 13x speed-up and 96% reduction in FLOPs. We collect an in-house edge and motion vector dataset with 16 types of actions by 40 users using the proprietary MVS. Our method outperforms individual modalities using only edge or motion vector images. Finally, we validate the privacy-enhanced quality of our sparse representation through face recognition on CelebA (a large face dataset) and a user study on our in-house dataset.
Code (0)
등록된 구현이 없습니다.
Tasks
Face RecognitionPose EstimationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Efficient Federated Learning with Enhanced Privacy via Lottery Ticket Pruning in Edge Computing
Federated learning (FL) is a collaborative learning paradigm for decentralized private data from mobile terminals (MTs). However, it suffers from issues in terms of communication, resource of MTs, and privacy. Existing p…
Edge-computingFederated LearningPrivacy PreservingPaired-Point Lifting for Enhanced Privacy-Preserving Visual Localization
Visual localization refers to the process of recovering camera pose from input image relative to a known scene, forming a cornerstone of numerous vision and robotics systems. While many algorithms utilize sparse 3D p…
feature selectionPrivacy PreservingVisual LocalizationEnhanced Sparse Point Cloud Data Processing for Privacy-aware Human Action Recognition
Human Action Recognition (HAR) plays a crucial role in healthcare, fitness tracking, and ambient assisted living technologies. While traditional vision based HAR systems are effective, they pose privacy concerns. mmWave …
Action RecognitionAre LLM-Enhanced GNNs Privacy-Safe?
Large language models (LLMs) have recently advanced graph neural networks (GNNs) by enriching node representations with semantic information, giving rise to LLM-enhanced GNNs that achieve substantial performance gains. H…
Graph LearningAI-Enhanced 3D RF Representation Using Low-Cost mmWave Radar
This paper introduces a system that takes radio frequency (RF) signals from an off-the-shelf, low-cost, 77 GHz mm Wave radar and produces an enhanced 3D RF representation of a scene. Such a system can be used in scenario…
RF-based Pose EstimationRobot Navigation