Transferable Feature Representation for Visible-to-Infrared Cross-Dataset Human Action Recognition
Recently, infrared human action recognition has attracted increasing attention for it has many advantages over visible light, that is, being robust to illumination change and shadows. However, the infrared action data is limited until now, which degrades the performance of infrared action recognition. Motivated by the idea of transfer learning, an infrared human action recognition framework using auxiliary data from visible light is proposed to solve the problem of limited infrared action data. In the proposed framework, we first construct a novel Cross-Dataset Feature Alignment and Generalization (CDFAG) framework to map the infrared data and visible light data into a common feature space, where Kernel Manifold Alignment (KEMA) and a dual alignedto-generalized encoders (AGE) model are employed to represent the feature. Then, a support vector machine (SVM) is trained, using both the infrared data and visible light data, and can classify the features derived from infrared data. The proposed method is evaluated on InfAR, which is a publicly available infrared human action dataset. To build up auxiliary data, we set up a novel visible light action dataset XD145. Experimental results show that the proposed method can achieve state-of-the-art performance compared with several transfer learning and domain adaptation methods.
Code (0)
등록된 구현이 없습니다.
Tasks
Action RecognitionDomain AdaptationTemporal Action LocalizationTransfer LearningSimilar Papers 제목 키워드 기반
Modality-Transition Representation Learning for Visible-Infrared Person Re-Identification
Visible-infrared person re-identification (VI-ReID) technique could associate the pedestrian images across visible and infrared modalities in the practical scenarios of background illumination changes. However, a substan…
Person Re-IdentificationRepresentation LearningHomogeneous and Heterogeneous Relational Graph for Visible-infrared Person Re-identification
Visible-infrared person re-identification (VI Re-ID) aims to match person images between the visible and infrared modalities. Existing VI Re-ID methods mainly focus on extracting homogeneous structural relationships in a…
Person Re-IdentificationLearning by Aligning: Visible-Infrared Person Re-identification using Cross-Modal Correspondences
We address the problem of visible-infrared person re-identification (VI-reID), that is, retrieving a set of person images, captured by visible or infrared cameras, in a cross-modal setting. Two main challenges in VI-reID…
Person Re-IdentificationRepresentation LearningCross-Modal Spherical Aggregation for Weakly Supervised Remote Sensing Shadow Removal
Remote sensing shadow removal, which aims to recover contaminated surface information, is tricky since shadows typically display overwhelmingly low illumination intensities. In contrast, the infrared image is robust towa…
Shadow RemovalFD2-Net: Frequency-Driven Feature Decomposition Network for Infrared-Visible Object Detection
Infrared-visible object detection (IVOD) seeks to harness the complementary information in infrared and visible images, thereby enhancing the performance of detectors in complex environments. However, existing methods of…
object-detectionObject Detection