paper-with-me

홈 › Papers

CAMO-MOT: Combined Appearance-Motion Optimization for 3D Multi-Object Tracking with Camera-LiDAR Fusion

2022-09-06 · Li Wang, Xinyu Zhang, Wenyuan Qin, Xiaoyu Li, Lei Yang, Zhiwei Li, Lei Zhu, Hong Wang, Jun Li, Huaping Liu

3D Multi-object tracking (MOT) ensures consistency during continuous dynamic detection, conducive to subsequent motion planning and navigation tasks in autonomous driving. However, camera-based methods suffer in the case of occlusions and it can be challenging to accurately track the irregular motion of objects for LiDAR-based methods. Some fusion methods work well but do not consider the untrustworthy issue of appearance features under occlusion. At the same time, the false detection problem also significantly affects tracking. As such, we propose a novel camera-LiDAR fusion 3D MOT framework based on the Combined Appearance-Motion Optimization (CAMO-MOT), which uses both camera and LiDAR data and significantly reduces tracking failures caused by occlusion and false detection. For occlusion problems, we are the first to propose an occlusion head to select the best object appearance features multiple times effectively, reducing the influence of occlusions. To decrease the impact of false detection in tracking, we design a motion cost matrix based on confidence scores which improve the positioning and object prediction accuracy in 3D space. As existing multi-object tracking methods only consider a single category, we also propose to build a multi-category loss to implement multi-object tracking in multi-category scenes. A series of validation experiments are conducted on the KITTI and nuScenes tracking benchmarks. Our proposed method achieves state-of-the-art performance and the lowest identity switches (IDS) value (23 for Car and 137 for Pedestrian) among all multi-modal MOT methods on the KITTI test dataset. And our proposed method achieves state-of-the-art performance among all algorithms on the nuScenes test dataset with 75.3% AMOTA.

📄 PDF Abstract BibTeX arXiv:2209.02540

Code (0)

등록된 구현이 없습니다.

Tasks

3D Multi-Object TrackingAutonomous DrivingMotion PlanningMulti-Object TrackingObjectObject Tracking

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

CamoSAM2: Motion-Appearance Induced Auto-Refining Prompts for Video Camouflaged Object Detection

2025-04-01 · Xin Zhang, Keren Fu, Qijun Zhao

The Segment Anything Model 2 (SAM2), a prompt-guided video foundation model, has remarkably performed in video object segmentation, drawing significant attention in the community. Due to the high similarity between camou…

Camouflaged Object Segmentationobject-detectionObject DetectionSemantic Segmentation+2

Location-Free Camouflage Generation Network

2022-03-18 · Yangyang Li, Wei Zhai, Yang Cao, Zheng-Jun Zha

Camouflage is a common visual phenomenon, which refers to hiding the foreground objects into the background images, making them briefly invisible to the human eye. Previous work has typically been implemented by an itera…

Mamba-based Spatio-Frequency Motion Perception for Video Camouflaged Object Detection

2025-07-31 · Xin Li, Keren Fu, Qijun Zhao arxiv

Existing video camouflaged object detection (VCOD) methods primarily rely on spatial appearances for motion perception. However, the high foreground-background similarity in VCOD limits the discriminability of such featu…

Temporal SequencesObject Detection

CAMotion: A High-Quality Benchmark for Camouflaged Moving Object Detection in the Wild

2026-04-09 · Siyuan Yao, Hao Sun, Ruiqi Yu, Xiwei Jiang 외 arxiv

Discovering camouflaged objects is a challenging task in computer vision due to the high similarity between camouflaged objects and their surroundings. While the problem of camouflaged object detection over sequential vi…

Moving Object Detection

Scoring, Remember, and Reference: Catching Camouflaged Objects in Videos

2025-03-21 · Yuang Feng, Shuyong Gao, Fuzhen Yan, Yicheng Song 외

Video Camouflaged Object Detection (VCOD) aims to segment objects whose appearances closely resemble their surroundings, posing a challenging and emerging task. Existing vision models often struggle in such scenarios due…

object-detectionObject Detection