paper-with-me

홈 › Papers

Prototype-based Cross-Modal Object Tracking

2023-12-22 · Lei Liu, Chenglong Li, Futian Wang, Longfeng Shen, Jin Tang

Cross-modal object tracking is an important research topic in the field of information fusion, and it aims to address imaging limitations in challenging scenarios by integrating switchable visible and near-infrared modalities. However, existing tracking methods face some difficulties in adapting to significant target appearance variations in the presence of modality switch. For instance, model update based tracking methods struggle to maintain stable tracking results during modality switching, leading to error accumulation and model drift. Template based tracking methods solely rely on the template information from first frame and/or last frame, which lacks sufficient representation ability and poses challenges in handling significant target appearance changes. To address this problem, we propose a prototype-based cross-modal object tracker called ProtoTrack, which introduces a novel prototype learning scheme to adapt to significant target appearance variations, for cross-modal object tracking. In particular, we design a multi-modal prototype to represent target information by multi-kind samples, including a fixed sample from the first frame and two representative samples from different modalities. Moreover, we develop a prototype generation algorithm based on two new modules to ensure the prototype representative in different challenges......

📄 PDF Abstract BibTeX arXiv:2312.14471

Code (1)

mmic-lcl/datasets-and-benchmark-code

Tasks

ObjectObject Tracking

Similar Papers 제목 키워드 기반

Prototypical Cross-Attention Networks for Multiple Object Tracking and Segmentation

2021-06-22 · NeurIPS 2021 12 · Lei Ke, Xia Li, Martin Danelljan, Yu-Wing Tai 외

Multiple object tracking and segmentation requires detecting, tracking, and segmenting objects belonging to a set of given classes. Most approaches only exploit the temporal dimension to address the association problem, …

Multi-Object Tracking and SegmentationMultiple Object Track and SegmentationMultiple Object TrackingObject+3

A2VIS: Amodal-Aware Approach to Video Instance Segmentation

2024-12-02 · Minh Tran, Thang Pham, Winston Bounsavy, Tri Nguyen 외

Handling occlusion remains a significant challenge for video instance-level tasks like Multiple Object Tracking (MOT) and Video Instance Segmentation (VIS). In this paper, we propose a novel framework, Amodal-Aware Video…

Instance SegmentationMultiple Object TrackingObjectObject Tracking+3

Small Object Tracking in LiDAR Point Cloud: Learning the Target-awareness Prototype and Fine-grained Search Region

2024-01-24 · Shengjing Tian, Yinan Han, Xiuping Liu, Xiantong Zhao

Single Object Tracking in LiDAR point cloud is one of the most essential parts of environmental perception, in which small objects are inevitable in real-world scenarios and will bring a significant barrier to the accura…

Object Tracking

VCP-DCN: Beyond Visual Concealed Property via Depth Collaborative Network for Camouflaged Object Detection

2026-07-30 · Songsong Duan, Xi Yang, Nannan Wang arxiv

Camouflaged Object Detection (COD) aims to identify and segment camouflaged objects in complex environments, which are often concealed because their color and texture are similar to the background. Several existing COD m…

Contrastive LearningObject Detection

Prototype-Aware Multimodal Alignment for Open-Vocabulary Visual Grounding

2025-09-08 · Jiangnan Xie, Xiaolong Zheng, Liang Zheng arxiv

Visual Grounding (VG) aims to utilize given natural language queries to locate specific target objects within images. While current transformer-based approaches demonstrate strong localization performance in standard sce…

Natural Language QueriesVisual Grounding