paper-with-me

홈 › Papers

Vision-Language Assistant for Emotional Reactions to Risky Driving

2026-07-17 · Harine Choi, Eun Hak Lee, Zhengzhong Tu arxiv

This study introduces a vision-language pipeline that detects risky driving behaviors and generates emotionally expressive responses to support driver awareness and comfort. Although vision-language models have advanced perception and reasoning in autonomous driving, existing systems rarely consider the emotional dimension or real-world user experience. Keep Yelling Assistant (KYA) detects high-risk driving maneuvers in real time, such as sudden cut-ins. It then produces emotional responses through a large language model tailored to driver preferences. The framework comprises two core modules. The vision module uses YOLOv8 variants to detect nearby vehicles and identify risky behaviors such as sudden cut-ins. Key driving metrics, including relative distance, speed, and projected reach time, are extracted and normalized to produce a structured behavior log. The language module processes this log with user-defined emotional tone settings, such as neutral, humorous, and analytical, and generates verbal reactions using state-of-the-art large language models, including ChatGPT-4o, Claude 3, Gemini 2.5, and Copilot. We evaluated the proposed system using dashcam videos containing risky driving behaviors and a user study involving 108 participants. Participants selected preferred response styles, and the large language models were evaluated based on emotional alignment. All models received favorable ratings, although preferences varied across personas. Notably, the combination of YOLOv8s and ChatGPT-4o achieved the highest score of 4.29 out of 5.00. By integrating real-world perception with emotionally adaptive dialogue, KYA introduces a new paradigm for emotionally intelligent in-vehicle artificial intelligence. It offers promising directions for improving safety, trust, and emotional well-being in both conventional and autonomous vehicles.

📄 PDF Abstract BibTeX arXiv:2607.16181

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous VehiclesAutonomous Driving

Similar Papers 제목 키워드 기반

Anatomy of a Feeling: Narrating Embodied Emotions via Large Vision-Language Models

2025-09-23 · Mohammad Saim, Phan Anh Duong, Cat Luong, Aniket Bhanderi 외 arxiv

The embodiment of emotional reactions from body parts contains rich information about our affective experiences. We propose a framework that utilizes state-of-the-art large vision-language models (LVLMs) to generate Embo…

Affection: Learning Affective Explanations for Real-World Visual Data

2022-10-04 · CVPR 2023 1 · Panos Achlioptas, Maks Ovsjanikov, Leonidas Guibas, Sergey Tulyakov

In this work, we explore the emotional reactions that real-world images tend to induce by using natural language as the medium to express the rationale behind an affective response to a given visual stimulus. To embark o…

Socratis: Are large multimodal models emotionally aware?

2023-08-31 · Katherine Deng, Arijit Ray, Reuben Tan, Saadia Gabriel 외

Existing emotion prediction benchmarks contain coarse emotion labels which do not consider the diversity of emotions that an image and text can elicit in humans due to various reasons. Learning diverse reactions to multi…

Articles

Incongruent Positivity: When Miscalibrated Positivity Undermines Online Supportive Conversations

2025-09-12 · Leen Almajed, Abeer ALdayel arxiv

In emotionally supportive conversations, well-intended positivity can sometimes misfire, leading to responses that feel dismissive, minimizing, or unrealistically optimistic. We examine this phenomenon of incongruent pos…

An Emotional Analysis of False Information in Social Media and News Articles

2019-08-26 · Bilal Ghanem, Paolo Rosso, Francisco Rangel

Fake news is risky since it has been created to manipulate the readers' opinions and beliefs. In this work, we compared the language of false news to the real one of real news from an emotional perspective, considering a…

Articles