paper-with-me

홈 › Papers

LaSOT: A High-quality Benchmark for Large-scale Single Object Tracking

2018-09-20 · CVPR 2019 6 · Heng Fan, Liting Lin, Fan Yang, Peng Chu, Ge Deng, Sijia Yu, Hexin Bai, Yong Xu, Chunyuan Liao, Haibin Ling

In this paper, we present LaSOT, a high-quality benchmark for Large-scale Single Object Tracking. LaSOT consists of 1,400 sequences with more than 3.5M frames in total. Each frame in these sequences is carefully and manually annotated with a bounding box, making LaSOT the largest, to the best of our knowledge, densely annotated tracking benchmark. The average video length of LaSOT is more than 2,500 frames, and each sequence comprises various challenges deriving from the wild where target objects may disappear and re-appear again in the view. By releasing LaSOT, we expect to provide the community with a large-scale dedicated benchmark with high quality for both the training of deep trackers and the veritable evaluation of tracking algorithms. Moreover, considering the close connections of visual appearance and natural language, we enrich LaSOT by providing additional language specification, aiming at encouraging the exploration of natural linguistic feature for tracking. A thorough experimental evaluation of 35 tracking algorithms on LaSOT is presented with detailed analysis, and the results demonstrate that there is still a big room for improvements.

📄 PDF Abstract BibTeX arXiv:1809.07845

Code (1)

HengLan/LaSOT_Evaluation_Toolkit

Tasks

Object TrackingVocal Bursts Intensity Prediction

Similar Papers 제목 키워드 기반

LaSOT: A High-quality Large-scale Single Object Tracking Benchmark

2020-09-08 · Heng Fan, Hexin Bai, Liting Lin, Fan Yang 외

Despite great recent advances in visual tracking, its further development, including both algorithm design and evaluation, is limited due to lack of dedicated large-scale benchmarks. To address this problem, we present L…

Object TrackingVisual TrackingVocal Bursts Intensity Prediction

Robust Visual Tracking by Motion Analyzing

2023-09-06 · Mohammed Leo, Kurban Ubul, ShengJie Cheng, Michael Ma

In recent years, Video Object Segmentation (VOS) has emerged as a complementary method to Video Object Tracking (VOT). VOS focuses on classifying all the pixels around the target, allowing for precise shape labeling, whi…

Object TrackingSegmentationSemantic SegmentationTensor Decomposition+4

Grounding-Tracking-Integration

2019-12-13 · Zhengyuan Yang, Tushar Kumar, Tianlang Chen, Jinsong Su 외

In this paper, we study Tracking by Language that localizes the target box sequence in a video based on a language query. We propose a framework called GTI that decomposes the problem into three sub-tasks: Grounding, Tra…

HiM2SAM: Enhancing SAM2 with Hierarchical Motion Estimation and Memory Optimization towards Long-term Tracking

2025-07-10 · Ruixiang Chen, Guolei Sun, Yawei Li, Jie Qin 외

This paper presents enhancements to the SAM2 framework for video object tracking task, addressing challenges such as occlusions, background clutter, and target reappearance. We introduce a hierarchical motion estimation …

Motion EstimationObject TrackingVideo Object Tracking

Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking

2025-06-25 · Ben Kang, Xin Chen, Jie Zhao, Chunjuan Bo 외

Transformer-based visual trackers have demonstrated significant advancements due to their powerful modeling capabilities. However, their practicality is limited on resource-constrained devices because of their slow proce…

GPUVisual Tracking