paper-with-me

홈 › Papers

STCT: Sequentially Training Convolutional Networks for Visual Tracking

2016-06-01 · CVPR 2016 6 · Lijun Wang, Wanli Ouyang, Xiaogang Wang, Huchuan Lu

Due to the limited amount of training samples, fine-tuning pre-trained deep models online is prone to over-fitting. In this paper, we propose a sequential training method for convolutional neural networks (CNNs) to effectively transfer pre-trained deep features for online applications. We regard a CNN as an ensemble with each channel of the output feature map as an individual base learner. Each base learner is trained using different loss criterions to reduce correlation and avoid over-training. To achieve the best ensemble online, all the base learners are sequentially sampled into the ensemble via important sampling. To further improve the robustness of each base learner, we propose to train the convolutional layers with random binary masks, which serves as a regularization to enforce each base learner to focus on different input features. The proposed online training method is applied to visual tracking problem by transferring deep features trained on massive annotated visual data and is shown to significantly improve tracking performance. Extensive experiments are conducted on two challenging benchmark data set and demonstrate that our tracking algorithm can outperform state-of-the-art methods with a considerable margin.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Tracking

Similar Papers 제목 키워드 기반

Set a Thief to Catch a Thief: Combating Label Noise through Noisy Meta Learning

2025-02-22 · Hanxuan Wang, Na Lu, Xueying Zhao, Yuxuan Yan 외

Learning from noisy labels (LNL) aims to train high-performance deep models using noisy datasets. Meta learning based label correction methods have demonstrated remarkable performance in LNL by designing various meta lab…

Meta-LearningRepresentation Learning

Analytical reconstructions of full-scan multiple source-translation computed tomography under large field of views

2023-05-31 · Zhisheng Wang, Yue Liu, Shunli Wang, Xingyuan Bian 외

This paper is to investigate the high-quality analytical reconstructions of multiple source-translation computed tomography (mSTCT) under an extended field of view (FOV). Under the larger FOVs, the previously proposed ba…

Intrapartum Ultrasound Image Segmentation of Pubic Symphysis and Fetal Head Using Dual Student-Teacher Framework with CNN-ViT Collaborative Learning

2024-09-11 · Jianmei Jiang, Huijin Wang, Jieyun Bai, Shun Long 외

The segmentation of the pubic symphysis and fetal head (PSFH) constitutes a pivotal step in monitoring labor progression and identifying potential delivery complications. Despite the advances in deep learning, the lack o…

Image SegmentationSegmentationSemantic Segmentation

BPF Algorithms for Multiple Source-Translation Computed Tomography Reconstruction

2023-05-30 · Zhisheng Wang, Haijun Yu, Yixing Huang, Shunli Wang 외

Micro-computed tomography (micro-CT) is a widely used state-of-the-art instrument employed to study the morphological structures of objects in various fields. However, its small field-of-view (FOV) cannot meet the pressi…

Translation

Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality Diagnosis

2026-02-01 · Haoran Lai, Zihang Jiang, Kun Zhang, Qingsong Yao 외 arxiv

Developing 3D vision-language models with robust clinical reasoning remains a challenge due to the inherent complexity of volumetric medical imaging, the tendency of models to overfit superficial report patterns, and the…

Visual Question AnsweringReinforcement Learning