paper-with-me

Papers

Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation

2025-02-06 · Namhee Kim, Woojin Park

Traditional autonomous driving systems often struggle with reasoning in complex, unexpected scenarios due to limited comprehension of spatial relationships. In response, this study introduces a Large Language Model (LLM)-based Autonomous Driving (AD) assistance system that integrates a vision adapter and an LLM reasoning module to enhance visual understanding and decision-making. The vision adapter, combining YOLOv4 and Vision Transformer (ViT), extracts comprehensive visual features, while GPT-4 enables human-like spatial reasoning and response generation. Experimental evaluations with 45 experienced drivers revealed that the system closely mirrors human performance in describing situations and moderately aligns with human decisions in generating appropriate responses.

📄 PDF Abstract BibTeX arXiv:2502.06843

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingDecision MakingLanguage ModelingLanguage ModellingLarge Language ModelResponse GenerationSpatial Reasoning

Methods 이 논문이 사용한 방법론

(TravEL!!Guide)How Do I File a Claim with Expedia? How Do I File a Claim with Expedia? Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Fast Help & Exclusive Travel Discounts!Need to file a claim with…
BNB Customer Service Number +1-833-534-1729 설명 없음
Attention 설명 없음
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Batch Normalization 설명 없음
Global Average Pooling Global Average Pooling is a pooling operation designed to replace fully connected layers in classical CNNs. The idea is to generate one feature map for each corresponding…
Tanh Activation 설명 없음
Average Pooling 설명 없음

Similar Papers 제목 키워드 기반

Vision-based Driver Assistance Systems: Survey, Taxonomy and Advances

2021-04-26 · Jonathan Horgan, Ciarán Hughes, John McDonald, Senthil Yogamani

Vision-based driver assistance systems is one of the rapidly growing research areas of ITS, due to various factors such as the increased level of safety requirements in automotive, computational power in embedded systems…

Autonomous DrivingSurvey

Testing Large Language Models on Driving Theory Knowledge and Skills for Connected Autonomous Vehicles

2024-07-24 · Zuoyin Tang, Jianhua He, Dashuai Pei, Kezhong Liu 외

Handling long tail corner cases is a major challenge faced by autonomous vehicles (AVs). While large language models (LLMs) hold great potentials to handle the corner cases with excellent generalization and explanation c…

Autonomous DrivingAutonomous Vehicles

A Survey on Multimodal Large Language Models for Autonomous Driving

2023-11-21 · Can Cui, Yunsheng Ma, Xu Cao, Wenqian Ye 외

With the emergence of Large Language Models (LLMs) and Vision Foundation Models (VFMs), multimodal AI systems benefiting from large models have the potential to equally perceive the real world, make decisions, and contro…

Autonomous Driving

AdaDrive: Self-Adaptive Slow-Fast System for Language-Grounded Autonomous Driving

2025-11-09 · Ruifei Zhang, Junlin Xie, Wei Zhang, Weikai Chen 외 arxiv

Effectively integrating Large Language Models (LLMs) into autonomous driving requires a balance between leveraging high-level reasoning and maintaining real-time efficiency. Existing approaches either activate LLMs too f…

Computational EfficiencyAutonomous Driving

Robustness of Segment Anything Model (SAM) for Autonomous Driving in Adverse Weather Conditions

2023-06-23 · Xinru Shan, Chaoning Zhang

Segment Anything Model (SAM) has gained considerable interest in recent times for its remarkable performance and has emerged as a foundational model in computer vision. It has been integrated in diverse downstream tasks,…

Autonomous Driving