paper-with-me

Papers

Multimodal-Enhanced Objectness Learner for Corner Case Detection in Autonomous Driving

2024-02-03 · Lixing Xiao, Ruixiao Shi, Xiaoyang Tang, Yi Zhou

Previous works on object detection have achieved high accuracy in closed-set scenarios, but their performance in open-world scenarios is not satisfactory. One of the challenging open-world problems is corner case detection in autonomous driving. Existing detectors struggle with these cases, relying heavily on visual appearance and exhibiting poor generalization ability. In this paper, we propose a solution by reducing the discrepancy between known and unknown classes and introduce a multimodal-enhanced objectness notion learner. Leveraging both vision-centric and image-text modalities, our semi-supervised learning framework imparts objectness knowledge to the student model, enabling class-aware detection. Our approach, Multimodal-Enhanced Objectness Learner (MENOL) for Corner Case Detection, significantly improves recall for novel classes with lower training costs. By achieving a 76.6% mAR-corner and 79.8% mAR-agnostic on the CODA-val dataset with just 5100 labeled training images, MENOL outperforms the baseline ORE by 71.3% and 60.6%, respectively. The code will be available at https://github.com/tryhiseyyysum/MENOL.

📄 PDF Abstract BibTeX arXiv:2402.02026

Code (1)

tryhiseyyysum/menol 공식 구현 pytorch

Tasks

Autonomous DrivingMultimodal Deep LearningObject Detection

Similar Papers 제목 키워드 기반

ECCV 2024 W-CODA: 1st Workshop on Multimodal Perception and Comprehension of Corner Cases in Autonomous Driving

2025-07-02 · Kai Chen, Ruiyuan Gao, Lanqing Hong, Hang Xu 외 arxiv

In this paper, we present details of the 1st W-CODA workshop, held in conjunction with the ECCV 2024. W-CODA aims to explore next-generation solutions for autonomous driving corner cases, empowered by state-of-the-art mu…

Scene UnderstandingAutonomous Driving

Adaptive Objectness for Object Tracking

2015-01-05 · Pengpeng Liang, Chunyuan Liao, Xue Mei, Haibin Ling

Object tracking is a long standing problem in vision. While great efforts have been spent to improve tracking performance, a simple yet reliable prior knowledge is left unexploited: the target object in tracking must be …

ObjectObject TrackingVisual Object TrackingVisual Tracking

Realistic Corner Case Generation for Autonomous Vehicles with Multimodal Large Language Model

2024-11-29 · QIUJING LU, Meng Ma, Ximiao Dai, Xuanhan Wang 외

To guarantee the safety and reliability of autonomous vehicle (AV) systems, corner cases play a crucial role in exploring the system's behavior under rare and challenging conditions within simulation environments. Howeve…

Autonomous VehiclesLanguage ModelingLanguage ModellingLarge Language Model+2

Memory-Bounded Left-Corner Unsupervised Grammar Induction on Child-Directed Input

2016-12-01 · COLING 2016 12 · Cory Shain, William Bryce, Lifeng Jin, Victoria Krakovna 외

This paper presents a new memory-bounded left-corner parsing model for unsupervised raw-text syntax induction, using unsupervised hierarchical hidden Markov models (UHHMM). We deploy this algorithm to shed light on the e…

Language AcquisitionSentence

Real-time Ultrasound-enhanced Multimodal Imaging of Tongue using 3D Printable Stabilizer System: A Deep Learning Approach

2019-11-22 · M. Hamed Mozaffari, Won-Sook Lee

Despite renewed awareness of the importance of articulation, it remains a challenge for instructors to handle the pronunciation needs of language learners. There are relatively scarce pedagogical tools for pronunciation …