paper-with-me

홈 › Papers

DRUformer: Enhancing the driving scene Important object detection with driving relationship self-understanding

2023-11-11 · Yingjie Niu, Ming Ding, Keisuke Fujii, Kento Ohtani, Alexander Carballo, Kazuya Takeda

Traffic accidents frequently lead to fatal injuries, contributing to over 50 million deaths until 2023. To mitigate driving hazards and ensure personal safety, it is crucial to assist vehicles in anticipating important objects during travel. Previous research on important object detection primarily assessed the importance of individual participants, treating them as independent entities and frequently overlooking the connections between these participants. Unfortunately, this approach has proven less effective in detecting important objects in complex scenarios. In response, we introduce Driving scene Relationship self-Understanding transformer (DRUformer), designed to enhance the important object detection task. The DRUformer is a transformer-based multi-modal important object detection model that takes into account the relationships between all the participants in the driving scenario. Recognizing that driving intention also significantly affects the detection of important objects during driving, we have incorporated a module for embedding driving intention. To assess the performance of our approach, we conducted a comparative experiment on the DRAMA dataset, pitting our model against other state-of-the-art (SOTA) models. The results demonstrated a noteworthy 16.2\% improvement in mIoU and a substantial 12.3\% boost in ACC compared to SOTA methods. Furthermore, we conducted a qualitative analysis of our model's ability to detect important objects across different road scenarios and classes, highlighting its effectiveness in diverse contexts. Finally, we conducted various ablation studies to assess the efficiency of the proposed modules in our DRUformer model.

📄 PDF Abstract BibTeX arXiv:2311.06497

Code (0)

등록된 구현이 없습니다.

Tasks

object-detectionObject Detection

Similar Papers 제목 키워드 기반

DrivingGaussian++: Towards Realistic Reconstruction and Editable Simulation for Surrounding Dynamic Driving Scenes

2025-08-28 · Yajiao Xiong, Xiaoyu Zhou, Yongtao Wan, Deqing Sun 외 arxiv

We present DrivingGaussian++, an efficient and effective framework for realistic reconstructing and controllable editing of surrounding dynamic autonomous driving scenes. DrivingGaussian++ models the static background us…

Autonomous Driving

Sce2DriveX: A Generalized MLLM Framework for Scene-to-Drive Learning

2025-02-19 · Rui Zhao, Qirui Yuan, Jinyu Li, Haofeng Hu 외

End-to-end autonomous driving, which directly maps raw sensor inputs to low-level vehicle controls, is an important part of Embodied AI. Despite successes in applying Multimodal Large Language Models (MLLMs) for high-lev…

Autonomous DrivingBench2DriveMotion PlanningQuestion Answering+3

How Do Drivers Allocate Their Potential Attention? Driving Fixation Prediction via Convolutional Neural Networks

2019-05-30 · IEEE Transactions on Intelligent Transportation Systems 2019 5 · Tao Deng, Hongmei Yan, Long Qin, Thuyen Ngo 외

The traffic driving environment is a complex and dynamic changing scene in which drivers have to pay close attention to salient and important targets or regions for safe driving. Modeling drivers’ eye movements and atten…

object-detectionObject Detection

Enhancing Traffic Scene Predictions with Generative Adversarial Networks

2019-09-24 · Peter König, Sandra Aigner, Marco Körner

We present a new two-stage pipeline for predicting frames of traffic scenes where relevant objects can still reliably be detected. Using a recent video prediction network, we first generate a sequence of future frames ba…

DeblurringImage Super-ResolutionImage-to-Image Translationobject-detection+6

BEV-LLM: Leveraging Multimodal BEV Maps for Scene Captioning in Autonomous Driving

2025-07-25 · Felix Brandstaetter, Erik Schuetz, Katharina Winter, Fabian Flohr arxiv

Autonomous driving technology has the potential to transform transportation, but its wide adoption depends on the development of interpretable and transparent decision-making systems. Scene captioning, which generates na…

Autonomous DrivingPoint Clouds