paper-with-me

홈 › Papers

PoseTrackReID: Dataset Description

2020-11-12 · Andreas Doering, Di Chen, Shanshan Zhang, Bernt Schiele, Juergen Gall

Current datasets for video-based person re-identification (re-ID) do not include structural knowledge in form of human pose annotations for the persons of interest. Nonetheless, pose information is very helpful to disentangle useful feature information from background or occlusion noise. Especially real-world scenarios, such as surveillance, contain a lot of occlusions in human crowds or by obstacles. On the other hand, video-based person re-ID can benefit other tasks such as multi-person pose tracking in terms of robust feature matching. For that reason, we present PoseTrackReID, a large-scale dataset for multi-person pose tracking and video-based person re-ID. With PoseTrackReID, we want to bridge the gap between person re-ID and multi-person pose tracking. Additionally, this dataset provides a good benchmark for current state-of-the-art methods on multi-frame person re-ID.

📄 PDF Abstract BibTeX arXiv:2011.06243

Code (0)

등록된 구현이 없습니다.

Tasks

Person Re-IdentificationPose TrackingVideo-Based Person Re-Identification

Similar Papers 제목 키워드 기반

Keypoint Message Passing for Video-based Person Re-Identification

2021-11-16 · Di Chen, Andreas Doering, Shanshan Zhang, Jian Yang 외

Video-based person re-identification (re-ID) is an important technique in visual surveillance systems which aims to match video snippets of people captured by different cameras. Existing methods are mostly based on convo…

Person Re-IdentificationRepresentation LearningVideo-Based Person Re-Identification

Cross-validating Image Description Datasets and Evaluation Metrics

2016-05-01 · LREC 2016 5 · Josiah Wang, Robert Gaizauskas

The task of automatically generating sentential descriptions of image content has become increasingly popular in recent years, resulting in the development of large-scale image description datasets and the proposal of va…

Image DescriptionSentence

Multi30K: Multilingual English-German Image Descriptions

2016-05-02 · WS 2016 8 · Desmond Elliott, Stella Frank, Khalil Sima'an, Lucia Specia

We introduce the Multi30K dataset to stimulate multilingual multimodal research. Recent advances in image description have been demonstrated on English-language datasets almost exclusively, but image description should n…

Image DescriptionMachine TranslationMultimodal Machine TranslationTranslation

TGIF: A New Dataset and Benchmark on Animated GIF Description

2016-04-10 · CVPR 2016 6 · Yuncheng Li, Yale Song, Liangliang Cao, Joel Tetreault 외

With the recent popularity of animated GIFs on social media, there is need for ways to index them with rich metadata. To advance research on animated GIF understanding, we collected a new dataset, Tumblr GIF (TGIF), with…

Image CaptioningMachine TranslationText GenerationTranslation+1

A Comprehensive Review on Recent Methods and Challenges of Video Description

2020-11-30 · Alok Singh, Thoudam Doren Singh, Sivaji Bandyopadhyay

Video description involves the generation of the natural language description of actions, events, and objects in the video. There are various applications of video description by filling the gap between languages and vis…

Machine TranslationSurveyVideo DescriptionVideo-Guided Machine Translation