paper-with-me

홈 › Papers

An Exploration of Target-Conditioned Segmentation Methods for Visual Object Trackers

2020-08-03 · Matteo Dunnhofer, Niki Martinel, Christian Micheloni

Visual object tracking is the problem of predicting a target object's state in a video. Generally, bounding-boxes have been used to represent states, and a surge of effort has been spent by the community to produce efficient causal algorithms capable of locating targets with such representations. As the field is moving towards binary segmentation masks to define objects more precisely, in this paper we propose to extensively explore target-conditioned segmentation methods available in the computer vision community, in order to transform any bounding-box tracker into a segmentation tracker. Our analysis shows that such methods allow trackers to compete with recently proposed segmentation trackers, while performing quasi real-time.

📄 PDF Abstract BibTeX arXiv:2008.00992

Code (0)

등록된 구현이 없습니다.

Tasks

Object TrackingSegmentationVisual Object Tracking

Similar Papers 제목 키워드 기반

Image-Conditioned Instance Prompt Network for Referring Remote Sensing Image Segmentation

2026-05-23 · Biaoyu Ren, Qingsheng Wang, Cun Xu, Dingkang Yang 외 arxiv

Referring Remote Sensing Image Segmentation (RRSIS) is a situated, task-driven cross-modal task related to the embodied perception paradigm, requiring models to align visual-spatial features with linguistic intentions fo…

Image Segmentation

HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios

2025-06-11 · Kunyu Peng, Junchao Huang, Xiangsheng Huang, Di Wen 외

Action segmentation is a core challenge in high-level video understanding, aiming to partition untrimmed videos into segments and assign each a label from a predefined action set. Existing methods primarily address singl…

Action RecognitionAction SegmentationBenchmarkingSegmentation+1

DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments

2026-05-15 · Anindya Sarkar, Srikumar Sastry, Aleksis Pirinen, Nathan Jacobs 외 arxiv

Visual active search (VAS) has been introduced as a modeling framework that leverages visual cues to direct aerial (e.g., UAV-based) exploration and pinpoint areas of interest within extensive geospatial regions. Potenti…

Reinforcement Learning

FAST-EQA: Efficient Embodied Question Answering with Global and Local Region Relevancy

2026-02-17 · Haochen Zhang, Nirav Savaliya, Faizan Siddiqui, Enna Sachdeva arxiv

Embodied Question Answering (EQA) combines visual scene understanding, goal-directed exploration, spatial and temporal reasoning under partial observability. A central challenge is to confine physical search to question-…

Scene UnderstandingQuestion Answering

Beyond One-to-One: Rethinking the Referring Image Segmentation

2023-08-26 · ICCV 2023 1 · Yutao Hu, Qixiong Wang, Wenqi Shao, Enze Xie 외

Referring image segmentation aims to segment the target object referred by a natural language expression. However, previous methods rely on the strong assumption that one sentence must describe one target in the image, w…

DecoderImage SegmentationImage to textSemantic Segmentation+1