paper-with-me

Papers

Context-aware Visual Tracking with Joint Meta-updating

2022-04-04 · Qiuhong Shen, Xin Li, Fanyang Meng, Yongsheng Liang

Visual object tracking acts as a pivotal component in various emerging video applications. Despite the numerous developments in visual tracking, existing deep trackers are still likely to fail when tracking against objects with dramatic variation. These deep trackers usually do not perform online update or update single sub-branch of the tracking model, for which they cannot adapt to the appearance variation of objects. Efficient updating methods are therefore crucial for tracking while previous meta-updater optimizes trackers directly over parameter space, which is prone to over-fit even collapse on longer sequences. To address these issues, we propose a context-aware tracking model to optimize the tracker over the representation space, which jointly meta-update both branches by exploiting information along the whole sequence, such that it can avoid the over-fitting problem. First, we note that the embedded features of the localization branch and the box-estimation branch, focusing on the local and global information of the target, are effective complements to each other. Based on this insight, we devise a context-aggregation module to fuse information in historical frames, followed by a context-aware module to learn affinity vectors for both branches of the tracker. Besides, we develop a dedicated meta-learning scheme, on account of fast and stable updating with limited training samples. The proposed tracking method achieves an EAO score of 0.514 on VOT2018 with the speed of 40FPS, demonstrating its capability of improving the accuracy and robustness of the underlying tracker with little speed drop.

📄 PDF Abstract BibTeX arXiv:2204.01513

Code (0)

등록된 구현이 없습니다.

Tasks

Meta-LearningObject TrackingVisual Object TrackingVisual Tracking

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Deep Meta Learning for Real-Time Target-Aware Visual Tracking

2017-12-26 · ICCV 2019 10 · Janghoon Choi, Junseok Kwon, Kyoung Mu Lee

In this paper, we propose a novel on-line visual tracking framework based on the Siamese matching network and meta-learner network, which run at real-time speeds. Conventional deep convolutional feature-based discriminat…

Meta-LearningReal-Time Visual TrackingVisual Tracking

Context-Aware Integration of Language and Visual References for Natural Language Tracking

2024-03-29 · CVPR 2024 1 · Yanyan Shao, Shuting He, Qi Ye, Yuchao Feng 외

Tracking by natural language specification (TNL) aims to consistently localize a target in a video sequence given a linguistic description in the initial frame. Existing methodologies perform language-based and template-…

Exploring Object Status Recognition for Recipe Progress Tracking in Non-Visual Cooking

2025-07-04 · Franklin Mingzhe Li, Kaitlyn Ng, Bin Zhu, Patrick Carrington arxiv

Cooking plays a vital role in everyday independence and well-being, yet remains challenging for people with vision impairments due to limited support for tracking progress and receiving contextual feedback. Object status…

Active Control Points-based 6DoF Pose Tracking for Industrial Metal Objects

2025-07-02 · Chentao Shen, Ding Pan, Mingyu Mei, Zaixing He 외 arxiv

Visual pose tracking is playing an increasingly vital role in industrial contexts in recent years. However, the pose tracking for industrial metal objects remains a challenging task especially in the real world-environme…

Pose Tracking

Predictive Visual Tracking: A New Benchmark and Baseline Approach

2021-03-08 · Bowen Li, Yiming Li, Junjie Ye, Changhong Fu 외

As a crucial robotic perception capability, visual tracking has been intensively studied recently. In the real-world scenarios, the onboard processing time of the image streams inevitably leads to a discrepancy between t…

Visual Tracking