paper-with-me

홈 › Papers

AccidentGPT: Large Multi-Modal Foundation Model for Traffic Accident Analysis

2024-01-05 · Kebin Wu, Wenbin Li, Xiaofei Xiao

Traffic accident analysis is pivotal for enhancing public safety and developing road regulations. Traditional approaches, although widely used, are often constrained by manual analysis processes, subjective decisions, uni-modal outputs, as well as privacy issues related to sensitive data. This paper introduces the idea of AccidentGPT, a foundation model of traffic accident analysis, which incorporates multi-modal input data to automatically reconstruct the accident process video with dynamics details, and furthermore provide multi-task analysis with multi-modal outputs. The design of the AccidentGPT is empowered with a multi-modality prompt with feedback for task-oriented adaptability, a hybrid training schema to leverage labelled and unlabelled data, and a edge-cloud split configuration for data privacy. To fully realize the functionalities of this model, we proposes several research opportunities. This paper serves as the stepping stone to fill the gaps in traditional approaches of traffic accident analysis and attract the research community attention for automatic, objective, and privacy-preserving traffic accident analysis.

📄 PDF Abstract BibTeX arXiv:2401.03040

Code (0)

등록된 구현이 없습니다.

Tasks

Privacy Preserving

Similar Papers 제목 키워드 기반

AccidentGPT: Accident Analysis and Prevention from V2X Environmental Perception with Multi-modal Large Model

2023-12-20 · Lening Wang, Yilong Ren, Han Jiang, Pinlong Cai 외

Traffic accidents, being a significant contributor to both human casualties and property damage, have long been a focal point of research for many scholars in the field of traffic safety. However, previous studies, wheth…

Autonomous DrivingScene Understanding

FOMO-3D: Using Vision Foundation Models for Long-Tailed 3D Object Detection

2026-03-09 · Anqi Joyce Yang, James Tu, Nikita Dvornik, Enxu Li 외 arxiv

In order to navigate complex traffic environments, self-driving vehicles must recognize many semantic classes pertaining to vulnerable road users or traffic control devices. However, many safety-critical objects (e.g., c…

3D Object Detection

Open-TransMind: A New Baseline and Benchmark for 1st Foundation Model Challenge of Intelligent Transportation

2023-04-12 · Yifeng Shi, Feng Lv, Xinliang Wang, Chunlong Xia 외

With the continuous improvement of computing power and deep learning algorithms in recent years, the foundation model has grown in popularity. Because of its powerful capabilities and excellent performance, this technolo…

2D Object DetectionImage RetrievalRetrieval

Towards Safe Mobility: A Unified Transportation Foundation Model enabled by Open-Ended Vision-Language Dataset

2026-04-24 · Wenhui Huang, Songyan Zhang, Collister Chua, Yang Liang 외 arxiv

Urban transportation systems face growing safety challenges that require scalable intelligence for emerging smart mobility infrastructures. While recent advances in foundation models and large-scale multimodal datasets h…

Visual Question AnsweringAutonomous Driving

MTP: Exploring Multimodal Urban Traffic Profiling with Modality Augmentation and Spectrum Fusion

2025-11-13 · Haolong Xiang, Peisi Wang, Xiaolong Xu, Kun Yi 외 arxiv

With rapid urbanization in the modern era, traffic signals from various sensors have been playing a significant role in monitoring the states of cities, which provides a strong foundation in ensuring safe travel, reducin…

Contrastive Learning