paper-with-me

Papers

Transformer Network for Multi-Person Tracking and Re-Identification in Unconstrained Environment

2023-12-19 · Hamza Mukhtar, Muhammad Usman Ghani Khan

Multi-object tracking (MOT) has profound applications in a variety of fields, including surveillance, sports analytics, self-driving, and cooperative robotics. Despite considerable advancements, existing MOT methodologies tend to falter when faced with non-uniform movements, occlusions, and appearance-reappearance scenarios of the objects. Recognizing this inadequacy, we put forward an integrated MOT method that not only marries object detection and identity linkage within a singular, end-to-end trainable framework but also equips the model with the ability to maintain object identity links over long periods of time. Our proposed model, named STMMOT, is built around four key modules: 1) candidate proposal generation, which generates object proposals via a vision-transformer encoder-decoder architecture that detects the object from each frame in the video; 2) scale variant pyramid, a progressive pyramid structure to learn the self-scale and cross-scale similarities in multi-scale feature maps; 3) spatio-temporal memory encoder, extracting the essential information from the memory associated with each object under tracking; and 4) spatio-temporal memory decoder, simultaneously resolving the tasks of object detection and identity association for MOT. Our system leverages a robust spatio-temporal memory module that retains extensive historical observations and effectively encodes them using an attention-based aggregator. The uniqueness of STMMOT lies in representing objects as dynamic query embeddings that are updated continuously, which enables the prediction of object states with attention mechanisms and eradicates the need for post-processing.

📄 PDF Abstract BibTeX arXiv:2312.11929

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderMulti-Object TrackingObjectobject-detectionObject DetectionObject TrackingSports Analytics

Similar Papers 제목 키워드 기반

Large-scale Multi-modal Person Identification in Real Unconstrained Environments

2019-12-17 · Jiajie Ye, Yisheng Guan, Junfa Liu, Xinghong Huang 외

Person identification (P-ID) under real unconstrained noisy environments is a huge challenge. In multiple-feature learning with Deep Convolutional Neural Networks (DCNNs) or Machine Learning method for large-scale person…

Multi-Modal Person IdentificationPerson Identificationvalid

PoseTrack: Joint Multi-Person Pose Estimation and Tracking

2016-11-23 · CVPR 2017 7 · Umar Iqbal, Anton Milan, Juergen Gall

In this work, we introduce the challenging problem of joint multi-person pose estimation and tracking of an unknown number of persons in unconstrained videos. Existing methods for multi-person pose estimation in images c…

Multi-Person Pose EstimationMulti-Person Pose Estimation and TrackingPose EstimationPose Tracking

A Gated Attention Transformer for Multi-Person Pose Tracking

2023-06-09 · Andreas Doering, Juergen Gall

Multi-person pose tracking is an important element for many applications and requires to estimate the human poses of all persons in a video and to track them over time. The association of poses across frames remains an o…

Pose Tracking

Unsupervised Noisy Tracklet Person Re-identification

2021-01-16 · Minxian Li, Xiatian Zhu, Shaogang Gong

Existing person re-identification (re-id) methods mostly rely on supervised model learning from a large set of person identity labelled training data per domain. This limits their scalability and usability in large scale…

One-Shot LearningPerson Re-Identification

Unconstrained Body Recognition at Altitude and Range: Comparing Four Approaches

2025-02-10 · Blake A Myers, Matthew Q Hill, Veda Nandan Gandi, Thomas M Metz 외

This study presents an investigation of four distinct approaches to long-term person identification using body shape. Unlike short-term re-identification systems that rely on temporary features (e.g., clothing), we focus…

Person Identification