paper-with-me

홈 › Papers

Precise Drive with VLM: First Prize Solution for PRCV 2024 Drive LM challenge

2024-11-05 · Bin Huang, Siyu Wang, Yuanpeng Chen, Yidan Wu, Hui Song, Zifan Ding, Jing Leng, Chengpeng Liang, Peng Xue, Junliang Zhang, Tiankun Zhao

This technical report outlines the methodologies we applied for the PRCV Challenge, focusing on cognition and decision-making in driving scenarios. We employed InternVL-2.0, a pioneering open-source multi-modal model, and enhanced it by refining both the model input and training methodologies. For the input data, we strategically concatenated and formatted the multi-view images. It is worth mentioning that we utilized the coordinates of the original images without transformation. In terms of model training, we initially pre-trained the model on publicly available autonomous driving scenario datasets to bolster its alignment capabilities of the challenge tasks, followed by fine-tuning on the DriveLM-nuscenes Dataset. During the fine-tuning phase, we innovatively modified the loss function to enhance the model's precision in predicting coordinate values. These approaches ensure that our model possesses advanced cognitive and decision-making capabilities in driving scenarios. Consequently, our model achieved a score of 0.6064, securing the first prize on the competition's final results.

📄 PDF Abstract BibTeX arXiv:2411.02999

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingDecision Making

Similar Papers 제목 키워드 기반

Infrared Small Target Detection based on Adjustable Sensitivity Strategy and Multi-Scale Fusion

2024-07-29 · Jinmiao Zhao, Zelin Shi, Chuang Yu, Yunpeng Liu

Recently, deep learning-based single-frame infrared small target (SIRST) detection technology has made significant progress. However, existing infrared small target detection methods are often optimized for a fixed image…

Sensitivity

Dataiku's Solution to SPHERE's Activity Recognition Challenge

2016-10-10 · Maxime Voisin, Leo Dreyfus-Schmidt, Pierre Gutierrez, Samuel Ronsin 외

Our team won the second prize of the Safe Aging with SPHERE Challenge organized by SPHERE, in conjunction with ECML-PKDD and Driven Data. The goal of the competition was to recognize activities performed by humans, using…

Activity RecognitionBIG-bench Machine LearningFeature Engineering

Sabotage and Free Riding in Contests with a Group-Specific Public-Good/Bad Prize

2025-02-12 · Kyung Hwan Baik, Dongwoo Lee

We study contests in which two groups compete to win (or not to win) a group-specific public-good/bad prize. Each player in the groups can exert two types of effort: one to help her own group win the prize, and one to sa…

Social Welfare in Search Games with Asymmetric Information

2020-06-26 · Gilad Bavly, Yuval Heller, Amnon Schreiber

We consider games in which players search for a hidden prize, and they have asymmetric information about the prize location. We study the social payoff in equilibria of these games. We present sufficient conditions for t…

The psychology of prizes: Loss aversion and optimal tournament rewards

2024-11-01 · Dmitry Ryvkin, Qin Wu

We study the optimal allocation of prizes in rank-order tournaments with loss averse agents. Prize sharing becomes increasingly optimal with loss aversion because more equitable prizes reduce the marginal psychological c…