paper-with-me

홈 › Papers

DecoderTracker: Decoder-Only Method for Multiple-Object Tracking

2023-10-26 · Liao Pan, Yang Feng, Wu Di, Liu Bo, Zhang Xingle

Decoder-only models, such as GPT, have demonstrated superior performance in many areas compared to traditional encoder-decoder structure transformer models. Over the years, end-to-end models based on the traditional transformer structure, like MOTR, have achieved remarkable performance in multi-object tracking. However, the significant computational resource consumption of these models leads to less friendly inference speeds and training times. To address these issues, this paper attempts to construct a lightweight Decoder-only model: DecoderTracker for end-to-end multi-object tracking. Specifically, drawing on some real-time detection models, we have developed an image feature extraction network which can efficiently extract features from images to replace the encoder structure. In addition to minor innovations in the network, we analyze the potential reasons for the slow training of MOTR-like models and propose an effective training strategy to mitigate the issue of prolonged training times. On the DanceTrack dataset, without any bells and whistles, DecoderTracker's tracking performance slightly surpasses that of MOTR, with approximately twice the inference speed. Furthermore, DecoderTracker requires significantly less training time compared to MOTR.

📄 PDF Abstract BibTeX arXiv:2310.17170

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderMulti-Object TrackingMultiple Object TrackingObjectobject-detectionObject DetectionObject Tracking

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Multi-Head Attention 설명 없음
Attention 설명 없음
Discriminative Fine-Tuning Discriminative Fine-Tuning is a fine-tuning strategy that is used for ULMFiT type models. Instead of using the same learning rate…
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Weight Decay 설명 없음
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Linear Warmup With Cosine Annealing Linear Warmup With Cosine Annealing is a learning rate schedule where we increase the learning rate linearly for $n$ updates and then anneal according to a cosine schedule…

Similar Papers 제목 키워드 기반

PatchTrack: Multiple Object Tracking Using Frame Patches

2022-01-01 · Xiaotong Chen, Seyed Mehdi Iranmanesh, Kuo-Chin Lien

Object motion and object appearance are commonly used information in multiple object tracking (MOT) applications, either for associating detections across frames in tracking-by-detection methods or direct track predictio…

DecoderMultiple Object TrackingObjectObject Tracking

TransMOT: Spatial-Temporal Graph Transformer for Multiple Object Tracking

2021-04-01 · Peng Chu, Jiang Wang, Quanzeng You, Haibin Ling 외

Tracking multiple objects in videos relies on modeling the spatial-temporal interactions of the objects. In this paper, we propose a solution named TransMOT, which leverages powerful graph transformers to efficiently mod…

DecoderMulti-Object TrackingMultiple Object TrackingObject+2

Joint Spatial-Temporal and Appearance Modeling with Transformer for Multiple Object Tracking

2022-05-31 · Peng Dai, Yiqiang Feng, Renliang Weng, ChangShui Zhang

The recent trend in multiple object tracking (MOT) is heading towards leveraging deep learning to boost the tracking performance. In this paper, we propose a novel solution named TransSTAM, which leverages Transformer to…

DecoderMultiple Object TrackingObject Tracking

TrackSSM: A General Motion Predictor by State-Space Model

2024-08-31 · Bin Hu, Run Luo, Zelin Liu, Cheng Wang 외

Temporal motion modeling has always been a key component in multiple object tracking (MOT) which can ensure smooth trajectory movement and provide accurate positional information to enhance association precision. However…

DecoderMambaMulti-Object TrackingMultiple Object Tracking+4

Jointly Optimizing State Operation Prediction and Value Generation for Dialogue State Tracking

2020-10-24 · Yan Zeng, Jian-Yun Nie

We investigate the problem of multi-domain Dialogue State Tracking (DST) with open vocabulary. Existing approaches exploit BERT encoder and copy-based RNN decoder, where the encoder predicts the state operation, and the …

DecoderDialogue State TrackingMulti-domain Dialogue State Tracking