paper-with-me

홈 › Papers

CLIPVehicle: A Unified Framework for Vision-based Vehicle Search

2025-08-06 · Likai Wang, Ruize Han, Xiangqun Zhang, Wei Feng arxiv

Vehicles, as one of the most common and significant objects in the real world, the researches on which using computer vision technologies have made remarkable progress, such as vehicle detection, vehicle re-identification, etc. To search an interested vehicle from the surveillance videos, existing methods first pre-detect and store all vehicle patches, and then apply vehicle re-identification models, which is resource-intensive and not very practical. In this work, we aim to achieve the joint detection and re-identification for vehicle search. However, the conflicting objectives between detection that focuses on shared vehicle commonness and re-identification that focuses on individual vehicle uniqueness make it challenging for a model to learn in an end-to-end system. For this problem, we propose a new unified framework, namely CLIPVehicle, which contains a dual-granularity semantic-region alignment module to leverage the VLMs (Vision-Language Models) for vehicle discrimination modeling, and a multi-level vehicle identification learning strategy to learn the identity representation from global, instance and feature levels. We also construct a new benchmark, including a real-world dataset CityFlowVS, and two synthetic datasets SynVS-Day and SynVS-All, for vehicle search. Extensive experimental results demonstrate that our method outperforms the state-of-the-art methods of both vehicle Re-ID and person search tasks.

📄 PDF Abstract BibTeX arXiv:2508.04120

Code (0)

등록된 구현이 없습니다.

Tasks

Vehicle Re-IdentificationPerson Search

Similar Papers 제목 키워드 기반

HM-ViT: Hetero-modal Vehicle-to-Vehicle Cooperative perception with vision transformer

2023-04-20 · ICCV 2023 1 · Hao Xiang, Runsheng Xu, Jiaqi Ma

Vehicle-to-Vehicle technologies have enabled autonomous vehicles to share information to see through occlusions, greatly enhancing perception performance. Nevertheless, existing works all focused on homogeneous traffic w…

Autonomous Vehicles

Neural Sentinel: Unified Vision Language Model (VLM) for License Plate Recognition with Human-in-the-Loop Continual Learning

2026-02-04 · Karthik Sivakoti arxiv

Traditional Automatic License Plate Recognition (ALPR) systems employ multi-stage pipelines consisting of object detection networks followed by separate Optical Character Recognition (OCR) modules, introducing compoundin…

License Plate RecognitionZero-shot GeneralizationAttribute ExtractionContinual Learning

To New Beginnings: A Survey of Unified Perception in Autonomous Vehicle Software

2025-08-28 · Loïc Stratil, Felix Fent, Esteban Rivera, Markus Lienkamp arxiv

Autonomous vehicle perception typically relies on modular pipelines that decompose the task into detection, tracking, and prediction. While interpretable, these pipelines suffer from error accumulation and limited inter-…

4C: A Computation, Communication, and Control Co-Design Framework for CAVs

2021-07-02 · Liangkai Liu, Shaoshan Liu, Weisong Shi

Connected and autonomous vehicles (CAVs) are promising due to their potential safety and efficiency benefits and have attracted massive investment and interest from government agencies, industry, and academia. With more …

Autonomous DrivingAutonomous Vehicles

V2X-ViT: Vehicle-to-Everything Cooperative Perception with Vision Transformer

2022-03-20 · Runsheng Xu, Hao Xiang, Zhengzhong Tu, Xin Xia 외

In this paper, we investigate the application of Vehicle-to-Everything (V2X) communication to improve the perception performance of autonomous vehicles. We present a robust cooperative perception framework with V2X commu…

3D Object DetectionAutonomous Vehiclesobject-detectionObject Detection