Unsupervised Multiple Person Tracking using AutoEncoder-Based Lifted Multicuts
Multiple Object Tracking (MOT) is a long-standing task in computer vision. Current approaches based on the tracking by detection paradigm either require some sort of domain knowledge or supervision to associate data correctly into tracks. In this work, we present an unsupervised multiple object tracking approach based on visual features and minimum cost lifted multicuts. Our method is based on straight-forward spatio-temporal cues that can be extracted from neighboring frames in an image sequences without superivison. Clustering based on these cues enables us to learn the required appearance invariances for the tracking task at hand and train an autoencoder to generate suitable latent representation. Thus, the resulting latent representations can serve as robust appearance cues for tracking even over large temporal distances where no reliable spatio-temporal features could be extracted. We show that, despite being trained without using the provided annotations, our model provides competitive results on the challenging MOT Benchmark for pedestrian tracking.
Code (0)
등록된 구현이 없습니다.
Tasks
ClusteringMultiple Object TrackingObject TrackingMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Lifted Disjoint Paths with Application in Multiple Object Tracking
We present an extension to the disjoint paths problem in which additional \emph{lifted} edges are introduced to provide path connectivity priors. We call the resulting optimization problem the lifted disjoint paths probl…
global-optimizationMultiple Object TrackingObjectObject TrackingMultiple People Tracking by Lifted Multicut and Person Re-Identification
Tracking multiple persons in a monocular video of a crowded scene is a challenging task. Humans can master it even if they loose track of a person locally by re-identifying the same person based on their appearance. Care…
Multiple People TrackingPerson Re-IdentificationPose EstimationLMGP: Lifted Multicut Meets Geometry Projections for Multi-Camera Multi-Object Tracking
Multi-Camera Multi-Object Tracking is currently drawing attention in the computer vision field due to its superior performance in real-world applications such as video surveillance in crowded scenes or in wide spaces. In…
3D geometryMulti-Object TrackingMultiple Object TrackingObject TrackingUnsupervised Spatio-temporal Latent Feature Clustering for Multiple-object Tracking and Segmentation
Assigning consistent temporal identifiers to multiple moving objects in a video sequence is a challenging problem. A solution to that problem would have immediate ramifications in multiple object tracking and segmentatio…
ClusteringInstance SegmentationMultiple Object TrackingObject Tracking+2Unsupervised Multiple-Object Tracking with a Dynamical Variational Autoencoder
In this paper, we present an unsupervised probabilistic model and associated estimation algorithm for multi-object tracking (MOT) based on a dynamical variational autoencoder (DVAE), called DVAE-UMOT. The DVAE is a laten…
Multi-Object TrackingMultiple Object TrackingObjectObject Tracking+2