PoseTrackReID: Dataset Description
Current datasets for video-based person re-identification (re-ID) do not include structural knowledge in form of human pose annotations for the persons of interest. Nonetheless, pose information is very helpful to disentangle useful feature information from background or occlusion noise. Especially real-world scenarios, such as surveillance, contain a lot of occlusions in human crowds or by obstacles. On the other hand, video-based person re-ID can benefit other tasks such as multi-person pose tracking in terms of robust feature matching. For that reason, we present PoseTrackReID, a large-scale dataset for multi-person pose tracking and video-based person re-ID. With PoseTrackReID, we want to bridge the gap between person re-ID and multi-person pose tracking. Additionally, this dataset provides a good benchmark for current state-of-the-art methods on multi-frame person re-ID.
Code (0)
등록된 구현이 없습니다.
Tasks
Person Re-IdentificationPose TrackingVideo-Based Person Re-IdentificationSimilar Papers 제목 키워드 기반
Keypoint Message Passing for Video-based Person Re-Identification
Video-based person re-identification (re-ID) is an important technique in visual surveillance systems which aims to match video snippets of people captured by different cameras. Existing methods are mostly based on convo…
Person Re-IdentificationRepresentation LearningVideo-Based Person Re-IdentificationCross-validating Image Description Datasets and Evaluation Metrics
The task of automatically generating sentential descriptions of image content has become increasingly popular in recent years, resulting in the development of large-scale image description datasets and the proposal of va…
Image DescriptionSentenceMulti30K: Multilingual English-German Image Descriptions
We introduce the Multi30K dataset to stimulate multilingual multimodal research. Recent advances in image description have been demonstrated on English-language datasets almost exclusively, but image description should n…
Image DescriptionMachine TranslationMultimodal Machine TranslationTranslationTGIF: A New Dataset and Benchmark on Animated GIF Description
With the recent popularity of animated GIFs on social media, there is need for ways to index them with rich metadata. To advance research on animated GIF understanding, we collected a new dataset, Tumblr GIF (TGIF), with…
Image CaptioningMachine TranslationText GenerationTranslation+1A Comprehensive Review on Recent Methods and Challenges of Video Description
Video description involves the generation of the natural language description of actions, events, and objects in the video. There are various applications of video description by filling the gap between languages and vis…
Machine TranslationSurveyVideo DescriptionVideo-Guided Machine Translation