paper-with-me

Papers

Multi-Adapter RGBT Tracking

2019-07-17 · Chenglong Li, Andong Lu, Aihua Zheng, Zhengzheng Tu, Jin Tang

The task of RGBT tracking aims to take the complementary advantages from visible spectrum and thermal infrared data to achieve robust visual tracking, and receives more and more attention in recent years. Existing works focus on modality-specific information integration by introducing modality weights to achieve adaptive fusion or learning robust feature representations of different modalities. Although these methods could effectively deploy the modality-specific properties, they ignore the potential values of modality-shared cues as well as instance-aware information, which are crucial for effective fusion of different modalities in RGBT tracking. In this paper, we propose a novel Multi-Adapter convolutional Network (MANet) to jointly perform modality-shared, modality-specific and instance-aware feature learning in an end-to-end trained deep framework for RGBT tracking. We design three kinds of adapters within our network. In a specific, the generality adapter is to extract shared object representations, the modality adapter aims at encoding modality-specific information to deploy their complementary advantages, and the instance adapter is to model the appearance properties and temporal variations of a certain object. Moreover, to reduce computational complexity for real-time demand of visual tracking, we design a parallel structure of generic adapter and modality adapter. Extensive experiments on two RGBT tracking benchmark datasets demonstrate the outstanding performance of the proposed tracker against other state-of-the-art RGB and RGBT tracking algorithms.

📄 PDF Abstract BibTeX arXiv:1907.07485

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Tracking

Similar Papers 제목 키워드 기반

RGBT Tracking via Multi-Adapter Network with Hierarchical Divergence Loss

2020-11-14 · Andong Lu, Chenglong Li, Yuqing Yan, Jin Tang 외

RGBT tracking has attracted increasing attention since RGB and thermal infrared data have strong complementary advantages, which could make trackers all-day and all-weather work. However, how to effectively represent RGB…

Representation LearningRgb-T TrackingVisual Tracking

UBATrack: Spatio-Temporal State Space Model for General Multi-Modal Tracking

2026-01-21 · Qihua Liang, Liang Chen, Yaozong Zheng, Jian Nong 외 arxiv

Multi-modal object tracking has attracted considerable attention by integrating multiple complementary inputs (e.g., thermal, depth, and event data) to achieve outstanding performance. Although current general-purpose mu…

Object Tracking

Breaking Shallow Limits: Task-Driven Pixel Fusion for Gap-free RGBT Tracking

2025-03-14 · Andong Lu, Yuanzhi Guo, Wanyu Wang, Chenglong Li 외

Current RGBT tracking methods often overlook the impact of fusion location on mitigating modality gap, which is key factor to effective tracking. Our analysis reveals that shallower fusion yields smaller distribution gap…

Representation LearningRgb-T Tracking

DRGBT-1K: A Large-scale High-quality Benchmark for Dynamic RGBT Tracking

2026-07-22 · Zhaodong Ding, Chenglong Li, Zeyu Ding, Futian Wang 외 arxiv

Dynamic RGBT (DRGBT) tracking aims to continuously localize a target when the available sensing modalities and observation platforms vary over time. Compared with conventional RGBT tracking with fixed RGBT inputs and a f…

LasHeR: A Large-scale High-diversity Benchmark for RGBT Tracking

2021-04-27 · Chenglong Li, Wanlin Xue, Yaqing Jia, Zhichen Qu 외

RGBT tracking receives a surge of interest in the computer vision community, but this research field lacks a large-scale and high-diversity benchmark dataset, which is essential for both the training of deep RGBT tracker…

DiversityRgb-T TrackingVocal Bursts Intensity Prediction