Leveraging Localization for Multi-camera Association
We present McAssoc, a deep learning approach to the as-sociation of detection bounding boxes in different views ofa multi-camera system. The vast majority of the academiahas been developing single-camera computer vision algo-rithms, however, little research attention has been directedto incorporating them into a multi-camera system. In thispaper, we designed a 3-branch architecture that leveragesdirect association and additional cross localization infor-mation. A new metric, image-pair association accuracy(IPAA) is designed specifically for performance evaluationof cross-camera detection association. We show in the ex-periments that localization information is critical to suc-cessful cross-camera association, especially when similar-looking objects are present. This paper is an experimentalwork prior to MessyTable, which is a large-scale bench-mark for instance association in mutliple cameras.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
CLIP-Loc: Multi-modal Landmark Association for Global Localization in Object-based Maps
This paper describes a multi-modal data association method for global localization using object-based maps and camera images. In global localization, or relocalization, using object-based maps, existing methods typically…
Language ModelingLanguage ModellingObjectCoarse-to-fine Semantic Localization with HD Map for Autonomous Driving in Structural Scenes
Robust and accurate localization is an essential component for robotic navigation and autonomous driving. The use of cameras for localization with high definition map (HD Map) provides an affordable localization sensor s…
Autonomous DrivingPose EstimationSemantic SegmentationGOReloc: Graph-based Object-Level Relocalization for Visual SLAM
This article introduces a novel method for object-level relocalization of robotic systems. It determines the pose of a camera sensor by robustly associating the object detections in the current frame with 3D objects in a…
Objectobject-detectionObject DetectionAHAP: Reconstructing Arbitrary Humans from Arbitrary Perspectives with Geometric Priors
Reconstructing 3D humans from images captured at multiple perspectives typically requires pre-calibration, like using checkerboards or MVS algorithms, which limits scalability and applicability in diverse real-world scen…
Camera Pose EstimationContrastive LearningMulti-HMR 2: Multi-Person Camera-Centric Human Detection, Mesh Recovery and Tracking
Most advances in human mesh recovery (HMR) have focused on pelvis-centered recovery, overlooking metric 3D localization and detection accuracy in the camera coordinate system - two key factors for real-world applications…
Human Mesh RecoveryScene Understanding