Scaling Human Activity Recognition: A Comparative Evaluation of Synthetic Data Generation and Augmentation Techniques
Human activity recognition (HAR) is often limited by the scarcity of labeled datasets due to the high cost and complexity of real-world data collection. To mitigate this, recent work has explored generating virtual inertial measurement unit (IMU) data via cross-modality transfer. While video-based and language-based pipelines have each shown promise, they differ in assumptions and computational cost. Moreover, their effectiveness relative to traditional sensor-level data augmentation remains unclear. In this paper, we present a direct comparison between these two virtual IMU generation approaches against classical data augmentation techniques. We construct a large-scale virtual IMU dataset spanning 100 diverse activities from Kinetics-400 and simulate sensor signals at 22 body locations. The three data generation strategies are evaluated on benchmark HAR datasets (UTD-MHAD, PAMAP2, HAD-AW) using four popular models. Results show that virtual IMU data significantly improves performance over real or augmented data alone, particularly under limited-data conditions. We offer practical guidance on choosing data generation strategies and highlight the distinct advantages and disadvantages of each approach.
Code (0)
등록된 구현이 없습니다.
Tasks
Activity RecognitionData AugmentationHuman Activity RecognitionSynthetic Data GenerationSimilar Papers 제목 키워드 기반
A Wearable Multi-Modal Edge-Computing System for Real-Time Kitchen Activity Recognition
In the human activity recognition research area, prior studies predominantly concentrate on leveraging advanced algorithms on public datasets to enhance recognition performance, little attention has been paid to executin…
Activity RecognitionEdge-computingHuman Activity RecognitionVision Language Models for Dynamic Human Activity Recognition in Healthcare Settings
As generative AI continues to evolve, Vision Language Models (VLMs) have emerged as promising tools in various healthcare applications. One area that remains relatively underexplored is their use in human activity recogn…
Human Activity RecognitionScaling laws in wearable human activity recognition
Many deep architectures and self-supervised pre-training techniques have been proposed for human activity recognition (HAR) from wearable multimodal sensors. Scaling laws have the potential to help move towards more prin…
Activity RecognitionHuman Activity RecognitionUncertainty-sensitive Activity Recognition: a Reliability Benchmark and the CARING Models
Beyond assigning the correct class, an activity recognition model should also be able to determine, how certain it is in its predictions. We present the first study of how welthe confidence values of modern action recogn…
Action RecognitionActivity Recognitionimage-classificationImage ClassificationEvaluationNet: Can Human Skill be Evaluated by Deep Networks?
With the recent substantial growth of media such as YouTube, a considerable number of instructional videos covering a wide variety of tasks are available online. Therefore, online instructional videos have become a rich …