paper-with-me

홈 › Papers

Towards Locally Deployable Fine-Tuned Causal Large Language Models for Mode Choice Behaviour

2025-07-29 · Tareq Alsaleh, Bilal Farooq arxiv

This study investigates the adoption of open-access, locally deployable causal large language models (LLMs) for travel mode choice prediction and introduces LiTransMC, the first fine-tuned causal LLM developed for this task. We systematically benchmark eleven open-access LLMs (1-12B parameters) across three stated and revealed preference datasets, testing 396 configurations and generating over 79,000 mode choice decisions. Beyond predictive accuracy, we evaluate models generated reasoning using BERTopic for topic modelling and a novel Explanation Strength Index, providing the first structured analysis of how LLMs articulate decision factors in alignment with behavioural theory. LiTransMC, fine-tuned using parameter efficient and loss masking strategy, achieved a weighted F1 score of 0.6845 and a Jensen-Shannon Divergence of 0.000245, surpassing both untuned local models and larger proprietary systems, including GPT-4o with advanced persona inference and embedding-based loading, while also outperforming classical mode choice methods such as discrete choice models and machine learning classifiers for the same dataset. This dual improvement, i.e., high instant-level accuracy and near-perfect distributional calibration, demonstrates the feasibility of creating specialist, locally deployable LLMs that integrate prediction and interpretability. Through combining structured behavioural prediction with natural language reasoning, this work unlocks the potential for conversational, multi-task transport models capable of supporting agent-based simulations, policy testing, and behavioural insight generation. These findings establish a pathway for transforming general purpose LLMs into specialized and explainable tools for transportation research and policy formulation, while maintaining privacy, reducing cost, and broadening access through local deployment.

📄 PDF Abstract BibTeX arXiv:2507.21432

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SkMTEB: Slovak Massive Text Embedding Benchmark and Model Adaptation

2026-06-11 · Marek Šuppa, Andrej Ridzik, Daniel Hládek, Natália Kňažeková 외 arxiv

We introduce SkMTEB, the first comprehensive MTEB-style text embedding benchmark for Slovak, a low-resource West Slavic language, comprising 31 datasets across 7 task types -- nearly 4$\times$ the depth of existing multi…

The Locally Deployable Virtual Doctor: LLM Based Human Interface for Automated Anamnesis and Database Conversion

2025-11-23 · Jan Benedikt Ruhland, Doguhan Bahcivan, Jan-Peter Sowa, Ali Canbay 외 arxiv

Recent advances in large language models made it possible to achieve high conversational performance with substantially reduced computational demands, enabling practical on-site deployment in clinical environments. Such …

Distilling Expert Surgical Knowledge: How to train local surgical VLMs for anatomy explanation in Complete Mesocolic Excision

2025-12-05 · Lennart Maack, Julia-Kristin Graß, Lisa-Marie Toscha, Nathaniel Melling 외 arxiv

Recently, Vision Large Language Models (VLMs) have demonstrated high potential in computer-aided diagnosis and decision-support. However, current VLMs show deficits in domain specific surgical scene understanding, such a…

Scene Understanding

ECG-LLM -- training and evaluation of domain-specific large language models for electrocardiography

2025-10-21 · Lara Ahrens, Wilhelm Haverkamp, Nils Strodthoff arxiv

Domain-adapted open-weight large language models (LLMs) offer promising healthcare applications, from queryable knowledge bases to multimodal assistants, with the crucial advantage of local deployment for privacy preserv…

CausalOPD: First-Wrong-Step Supervision for Distilling Causal Chain Reasoning

2026-08-04 · Jian Zhang, Bingyi Wang, Yizhi Liu arxiv

Many critical reasoning tasks, including clinical diagnosis, legal judgment, and industrial fault diagnosis, require step-dependent causal chains in which early errors propagate and correct conclusions can mask invalid r…

Reinforcement LearningFault Diagnosis