paper-with-me

Papers

Visual Object Tracking in First Person Vision

2022-09-27 · Matteo Dunnhofer, Antonino Furnari, Giovanni Maria Farinella, Christian Micheloni

The understanding of human-object interactions is fundamental in First Person Vision (FPV). Visual tracking algorithms which follow the objects manipulated by the camera wearer can provide useful information to effectively model such interactions. In the last years, the computer vision community has significantly improved the performance of tracking algorithms for a large variety of target objects and scenarios. Despite a few previous attempts to exploit trackers in the FPV domain, a methodical analysis of the performance of state-of-the-art trackers is still missing. This research gap raises the question of whether current solutions can be used ``off-the-shelf'' or more domain-specific investigations should be carried out. This paper aims to provide answers to such questions. We present the first systematic investigation of single object tracking in FPV. Our study extensively analyses the performance of 42 algorithms including generic object trackers and baseline FPV-specific trackers. The analysis is carried out by focusing on different aspects of the FPV setting, introducing new performance measures, and in relation to FPV-specific tasks. The study is made possible through the introduction of TREK-150, a novel benchmark dataset composed of 150 densely annotated video sequences. Our results show that object tracking in FPV poses new challenges to current visual trackers. We highlight the factors causing such behavior and point out possible research directions. Despite their difficulties, we prove that trackers bring benefits to FPV downstream tasks requiring short-term object tracking. We expect that generic object tracking will gain popularity in FPV as new and FPV-specific methodologies are investigated.

📄 PDF Abstract BibTeX arXiv:2209.13502

Code (0)

등록된 구현이 없습니다.

Tasks

Human-Object Interaction DetectionObjectObject TrackingVisual Object TrackingVisual Tracking

Similar Papers 제목 키워드 기반

Is First Person Vision Challenging for Object Tracking?

2021-08-31 · Matteo Dunnhofer, Antonino Furnari, Giovanni Maria Farinella, Christian Micheloni

Understanding human-object interactions is fundamental in First Person Vision (FPV). Tracking algorithms which follow the objects manipulated by the camera wearer can provide useful cues to effectively model such interac…

Human-Object Interaction DetectionObjectObject TrackingVisual Tracking

Is Tracking really more challenging in First Person Egocentric Vision?

2025-07-21 · Matteo Dunnhofer, Zaira Manigrasso, Christian Micheloni arxiv

Visual object tracking and segmentation are becoming fundamental tasks for understanding human activities in egocentric vision. Recent research has benchmarked state-of-the-art methods and concluded that first person ego…

Visual Object Tracking

Is First Person Vision Challenging for Object Tracking?

2020-11-24 · Matteo Dunnhofer, Antonino Furnari, Giovanni Maria Farinella, Christian Micheloni

Understanding human-object interactions is fundamental in First Person Vision (FPV). Tracking algorithms which follow the objects manipulated by the camera wearer can provide useful cues to effectively model such interac…

Human-Object Interaction DetectionObjectObject TrackingVisual Object Tracking+1

An On-line Variational Bayesian Model for Multi-Person Tracking from Cluttered Scenes

2015-09-04 · Sileye . Ba, Xavier Alameda-Pineda, Alessio Xompero, Radu Horaud

Object tracking is an ubiquitous problem that appears in many applications such as remote sensing, audio processing, computer vision, human-machine interfaces, human-robot interaction, etc. Although thoroughly investigat…

Multiple Object TrackingObjectObject Tracking

Teaching VLMs to Localize Specific Objects from In-context Examples

2024-11-20 · Sivan Doveh, Nimrod Shabtay, Wei Lin, Eli Schwartz 외

Vision-Language Models (VLMs) have shown remarkable capabilities across diverse visual tasks, including image recognition, video understanding, and Visual Question Answering (VQA) when explicitly trained for these tasks.…

ObjectObject TrackingQuestion AnsweringVideo Object Tracking+3