3D Shape Temporal Aggregation for Video-Based Clothing-Change Person Re-Identication
3D shape of human body can be both discriminative and clothing-independent information in video-based clothing-change person re-identification (Re-ID). However, existing Re-ID methods usually generate 3D body shapes without considering identity modeling, which severely weakens the discriminability of 3D human shapes. In addition, different video frames provide highly similar 3D shapes, but existing methods cannot capture the differences among 3D shapes over time. They are thus insensitive to the unique and discriminative 3D shape information of each frame and ineffectively aggregate many redundant framewise shapes in a video-wise representation for Re-ID. To address these problems, we propose a 3D Shape Temporal Aggregation (3STA) model for video-based clothing-change Re-ID. To generate the discriminative 3D shape for each frame, we rst introduce an identity-aware 3D shape generation module. It embeds the identity information into the generation of 3D shapes by the joint learning of shape estimation and identity recognition. Second, a difference-aware shape aggregation module is designed to measure inter-frame 3D human shape differences and automatically select the unique 3D shape information of each frame. This helps minimize redundancy and maximize complementarity in temporal shape aggregation. We further construct a Video-based Clothing-Change ReID (VCCR) dataset to address the lack of publicly available datasets for video-based clothing-change Re-ID. Extensive experiments on the VCCR dataset demonstrate the effectiveness of the proposed 3STA model. The dataset is available at https://vhank.github.io/vccr.github.io.
Code (1)
Tasks
3D Shape GenerationPerson Re-IdentificationSimilar Papers 제목 키워드 기반
Temporal 3D Shape Modeling for Video-Based Cloth-Changing Person Re-Identification
Video-based Cloth-Changing Person Re-ID (VCCRe-ID) refers to a real-world Re-ID problem where texture information like appearance or clothing becomes unreliable in long-term, limiting the applicability of traditional Re-…
3D Shape ModelingCloth-Changing Person Re-IdentificationPerson Re-IdentificationAttention-based Shape and Gait Representations Learning for Video-based Cloth-Changing Person Re-Identification
Current state-of-the-art Video-based Person Re-Identification (Re-ID) primarily relies on appearance features extracted by deep learning models. These methods are not applicable for long-term analysis in real-world scena…
Cloth-Changing Person Re-IdentificationGraph AttentionPerson Re-IdentificationVideo-Based Person Re-IdentificationMonoClothCap: Towards Temporally Coherent Clothing Capture from Monocular RGB Video
We present a method to capture temporally coherent dynamic clothing deformation from a monocular RGB video input. In contrast to the existing literature, our method does not require a pre-scanned personalized mesh templa…
Surface ReconstructionvalidDress like a Star: Retrieving Fashion Products from Videos
This work proposes a system for retrieving clothing and fashion products from video content. Although films and television are the perfect showcase for fashion brands to promote their products, spectators are not always …
RetrievalLearning Shape Representations for Clothing Variations in Person Re-Identification
Person re-identification (re-ID) aims to recognize instances of the same person contained in multiple images taken across different cameras. Existing methods for re-ID tend to rely heavily on the assumption that both que…
DisentanglementPerson Re-IdentificationRepresentation Learning