A baseline on continual learning methods for video action recognition
Continual learning has recently attracted attention from the research community, as it aims to solve long-standing limitations of classic supervisedly-trained models. However, most research on this subject has tackled continual learning in simple image classification scenarios. In this paper, we present a benchmark of state-of-the-art continual learning methods on video action recognition. Besides the increased complexity due to the temporal dimension, the video setting imposes stronger requirements on computing resources for top-performing rehearsal methods. To counteract the increased memory requirements, we present two method-agnostic variants for rehearsal methods, exploiting measures of either model confidence or data information to select memorable samples. Our experiments show that, as expected from the literature, rehearsal methods outperform other approaches; moreover, the proposed memory-efficient variants are shown to be effective at retaining a certain level of performance with a smaller buffer size.
Code (0)
등록된 구현이 없습니다.
Tasks
Action RecognitionContinual Learningimage-classificationImage ClassificationTemporal Action LocalizationSimilar Papers 제목 키워드 기반
Video Domain Incremental Learning for Human Action Recognition in Home Environments
It is significantly challenging to recognize daily human actions in homes due to the diversity and dynamic changes in unconstrained home environments. It spurs the need to continually adapt to various users and scenes. F…
Action Recognitionclass-incremental learningClass Incremental LearningContinual Learning+3Towards Continual Egocentric Activity Recognition: A Multi-modal Egocentric Activity Dataset for Continual Learning
With the rapid development of wearable cameras, a massive collection of egocentric video for first-person visual perception becomes available. Using egocentric videos to predict first-person activity faces many challenge…
Activity RecognitionContinual LearningEgocentric Activity RecognitionHuman Activity RecognitionClass-Incremental Learning for Action Recognition in Videos
We tackle catastrophic forgetting problem in the context of class-incremental learning for video recognition, which has not been explored actively despite the popularity of continual learning. Our framework addresses thi…
Action RecognitionAction Recognition In Videosclass-incremental learningClass Incremental Learning+4CM2-Net: Continual Cross-Modal Mapping Network for Driver Action Recognition
Driver action recognition has significantly advanced in enhancing driver-vehicle interactions and ensuring driving safety by integrating multiple modalities, such as infrared and depth. Nevertheless, compared to RGB moda…
Action RecognitionContinual LearningCLVOS23: A Long Video Object Segmentation Dataset for Continual Learning
Continual learning in real-world scenarios is a major challenge. A general continual learning model should have a constant memory size and no predefined task boundaries, as is the case in semi-supervised Video Object Seg…
Continual LearningSemantic SegmentationSemi-Supervised Video Object SegmentationVideo Object Segmentation+1