paper-with-me

Papers

DriveSafe: A Framework for Risk Detection and Safety Suggestions in Driving Scenarios

2026-05-16 · Sainithin Artham, Shankar Gangisetty, Avijit Dasgupta, C. V. Jawahar arxiv

Comprehensive situational awareness is essential for autonomous vehicles operating in safety-critical environments, as it enables the identification and mitigation of potential risks. Although recent Multimodal Large Language Models (MLLMs) have shown promise on general vision-language tasks, our findings indicate that zero-shot MLLMs still underperform compared to domain-specific methods in fine-grained, spatially grounded risk assessment. To address this gap, we propose DriveSafe, a framework for risk-aware scene understanding that leverages structured natural language descriptions. Specifically, our method first generates spatially grounded captions enriched with multimodal context, including motion, spatial, and depth cues. These captions are then used for downstream risk assessment, explicitly identifying hazardous objects, their locations, and the unsafe behaviors they imply, followed by actionable safety suggestions. To further improve performance, we employ caption-risk pairings to fine-tune a lightweight adapter module, efficiently injecting domain-specific knowledge into the base LLM. By conditioning risk assessment on explicit language-based scene representations, DriveSafe achieves significant gains over both zero-shot MLLMs and prior domain-specific baselines. Exhaustive experiments on the DRAMA benchmark demonstrate state-of-the-art performance, while ablation studies validate the effectiveness of our key design choices. Project page: https://cvit.iiit.ac.in/ research/projects/cvit-projects/drivesafe

📄 PDF Abstract BibTeX arXiv:2605.16892

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous VehiclesScene Understanding

Similar Papers 제목 키워드 기반

DriveSafe: A Hierarchical Risk Taxonomy for Safety-Critical LLM-Based Driving Assistants

2026-01-17 · Abhishek Kumar, Riya Tapwal, Carsten Maple arxiv

Large Language Models (LLMs) are increasingly integrated into vehicle-based digital assistants, where unsafe, ambiguous, or legally incorrect responses can lead to serious safety, ethical, and regulatory consequences. De…

DriveSafer: End-to-End Autonomous Driving with Safety Guidance

2026-05-16 · Shounak Sural, Raj Rajkumar arxiv

End-to-End (E2E) autonomous driving models have shown growing capability in recent years, with performance improving on increasingly challenging benchmarks. However, modern generative E2E planners still suffer from a sub…

Autonomous Driving

Multi-strategy Collaborative Optimized YOLOv5s and its Application in Distance Estimation

2023-12-07 · Zijian Shen, Zhenping Mu, Xiangxiang Li

The increasing accident rate brought about by the explosive growth of automobiles has made the research on active safety systems of automobiles increasingly important. The importance of improving the accuracy of vehicle …

vehicle detection

Beyond Reactive Safety: Risk-Aware LLM Alignment via Long-Horizon Simulation

2025-06-26 · Chenkai Sun, Denghui Zhang, ChengXiang Zhai, Heng Ji

Given the growing influence of language model-based agents on high-stakes societal decisions, from public policy to healthcare, ensuring their beneficial impact requires understanding the far-reaching implications of the…

Language ModelingLanguage Modelling

DriftGuard: Safety-Aware Multi-Monitor Detection and Selective Adaptation for Evolving Toxicity Moderation

2026-06-27 · Yuting Xin, Hanyu Cai, Binqi Shen, Lier Jin 외 arxiv

Automated toxicity moderation systems operate in dynamic online environments where harmful behavior evolves through coded language, shifting targets, and strategic adaptation to enforcement. Existing drift detection meth…