paper-with-me

Papers

ARGUSTRACK: A Multi-View Annotation System for Multi-Object Tracking

2026-06-14 · Hao Vo, Duc Nguyen, Ngan Le arxiv

Multi-Camera Multi-Target (MCMT) tracking has emerged as a critical capability for applications ranging from autonomous driving to animal behavior monitoring. While recent advances have yielded sophisticated tracking algorithms, the availability of annotated multi-view data remains a significant bottleneck. Existing annotation tools predominantly support single-camera workflows or rely on LiDAR sensors, making cross-view labeling tedious and impractical for camera-only setups. We present ARGUS-TRACK, a multi-camera annotation system that addresses these limitations by enabling annotators to work directly on a bird's-eye-view (BEV) plane. Given calibrated camera parameters, a single ground-plane annotation is automatically projected into 2D bounding boxes across all relevant views, inherently ensuring identity consistency without manual cross-view alignment. To further accelerate the labeling process, ARGUSTRACK incorporates two complementary mechanisms: a Temporal Aware module that propagates annotations from preceding frames to initialize new ones, requiring only minor positional adjustments; and a Multi-camera Semi-annotation module that leverages off-the-shelf 2D detectors combined with foot-point estimation to automatically generate candidate BEV positions for annotator verification. We evaluate ARGUSTRACK through a pilot study on multi-camera broiler tracking and demonstrate that it substantially reduces annotation time compared to conventional single-camera labeling workflows.

📄 PDF Abstract BibTeX arXiv:2606.20687

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-Object TrackingAutonomous Driving

Similar Papers 제목 키워드 기반

Rethinking the Data Annotation Process for Multi-view 3D Pose Estimation with Active Learning and Self-Training

2021-12-27 · Qi Feng, Kun He, He Wen, Cem Keskin 외

Pose estimation of the human body and hands is a fundamental problem in computer vision, and learning-based solutions require a large amount of annotated data. In this work, we improve the efficiency of the data annotati…

3D Pose EstimationActive LearningHand Pose EstimationPose Estimation

ViMRHP: A Vietnamese Benchmark Dataset for Multimodal Review Helpfulness Prediction via Human-AI Collaborative Annotation

2025-05-12 · Truc Mai-Thanh Nguyen, Dat Minh Nguyen, Son T. Luu, Kiet Van Nguyen

Multimodal Review Helpfulness Prediction (MRHP) is an essential task in recommender systems, particularly in E-commerce platforms. Determining the helpfulness of user-generated reviews enhances user experience and improv…

2kRecommendation Systems

A multimodal movie review corpus for fine-grained opinion mining

2019-02-26 · Alexandre Garcia, Slim Essid, Florence d'Alché-Buc, Chloé Clavel

In this paper, we introduce a set of opinion annotations for the POM movie review dataset, composed of 1000 videos. The annotation campaign is motivated by the development of a hierarchical opinion prediction framework a…

Opinion MiningPrediction

View-aware Cross-modal Distillation for Multi-view Action Recognition

2025-11-17 · Trung Thanh Nguyen, Yasutomo Kawanishi, Vijay John, Takahiro Komamizu 외 arxiv

The widespread use of multi-sensor systems has increased research in multi-view action recognition. While existing approaches in multi-view setups with fully overlapping sensors benefit from consistent view coverage, par…

Knowledge DistillationAction Recognition

Multiview-Consistent Semi-Supervised Learning for 3D Human Pose Estimation

2019-08-14 · CVPR 2020 6 · Rahul Mitra, Nitesh B. Gundavarapu, Abhishek Sharma, Arjun Jain

The best performing methods for 3D human pose estimation from monocular images require large amounts of in-the-wild 2D and controlled 3D pose annotated datasets which are costly and require sophisticated systems to acqui…

3D Human Pose EstimationMetric LearningPose EstimationPose Retrieval+1