paper-with-me

홈 › Papers

From Two-Stream to One-Stream: Efficient RGB-T Tracking via Mutual Prompt Learning and Knowledge Distillation

2024-03-25 · Yang Luo, Xiqing Guo, Hao Li

Due to the complementary nature of visible light and thermal infrared modalities, object tracking based on the fusion of visible light images and thermal images (referred to as RGB-T tracking) has received increasing attention from researchers in recent years. How to achieve more comprehensive fusion of information from the two modalities at a lower cost has been an issue that researchers have been exploring. Inspired by visual prompt learning, we designed a novel two-stream RGB-T tracking architecture based on cross-modal mutual prompt learning, and used this model as a teacher to guide a one-stream student model for rapid learning through knowledge distillation techniques. Extensive experiments have shown that, compared to similar RGB-T trackers, our designed teacher model achieved the highest precision rate, while the student model, with comparable precision rate to the teacher model, realized an inference speed more than three times faster than the teacher model.(Codes will be available if accepted.)

📄 PDF Abstract BibTeX arXiv:2403.16834

Code (0)

등록된 구현이 없습니다.

Tasks

Knowledge DistillationObject TrackingPrompt LearningRgb-T Tracking

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…
Knowledge Distillation A very simple way to improve the performance of almost any machine learning algorithm is to train many different models on the same data and then to average their predictions.…

Similar Papers 제목 키워드 기반

RGB-T Tracking via Multi-Modal Mutual Prompt Learning

2023-08-31 · Yang Luo, Xiqing Guo, Hui Feng, Lei Ao

Object tracking based on the fusion of visible and thermal im-ages, known as RGB-T tracking, has gained increasing atten-tion from researchers in recent years. How to achieve a more comprehensive fusion of information fr…

Object TrackingPrompt LearningRgb-T Tracking

Visual Prompt Multi-Modal Tracking

2023-03-20 · CVPR 2023 1 · Jiawen Zhu, Simiao Lai, Xin Chen, Dong Wang 외

Visible-modal object tracking gives rise to a series of downstream multi-modal tracking tributaries. To inherit the powerful representations of the foundation model, a natural modus operandi for multi-modal tracking is f…

Object TrackingPrompt LearningRgb-T Tracking

Middle Fusion and Multi-Stage, Multi-Form Prompts for Robust RGB-T Tracking

2024-03-27 · Qiming Wang, Yongqiang Bai, Hongxing Song

RGB-T tracking, a vital downstream task of object tracking, has made remarkable progress in recent years. Yet, it remains hindered by two major challenges: 1) the trade-off between performance and efficiency; 2) the scar…

FormObject TrackingPrompt LearningRgb-T Tracking

VL-UniTrack: A Unified Framework with Visual-Language Prompts for UAV-Ground Visual Tracking

2026-05-06 · Boyue Xu, Ruichao Hou, Tongwei Ren, Gangshan Wu arxiv

UAV-ground visual tracking (UGVT) aims to simultaneously track the same object from both the UAV and the ground view. However, existing two-stream methods suffer from isolated feature extraction and rely heavily on impli…

Visual Tracking

Fine-Grained Shape-Appearance Mutual Learning for Cloth-Changing Person Re-Identification

2021-06-19 · CVPR 2021 1 · Peixian Hong, Tao Wu, AnCong Wu, Xintong Han 외

Recently, person re-identification (Re-ID) has achieved great progress. However, current methods largely depend on color appearance, which is not reliable when a person changes the clothes. Cloth-changing Re-ID is ch…

Cloth-Changing Person Re-IdentificationPerson Re-Identification