paper-with-me

홈 › Papers

Using Reinforcement Learning to Train Large Language Models to Explain Human Decisions

2025-05-16 · Jian-Qiao Zhu, Hanbo Xie, Dilip Arumugam, Robert C. Wilson, Thomas L. Griffiths

A central goal of cognitive modeling is to develop models that not only predict human behavior but also provide insight into the underlying cognitive mechanisms. While neural network models trained on large-scale behavioral data often achieve strong predictive performance, they typically fall short in offering interpretable explanations of the cognitive processes they capture. In this work, we explore the potential of pretrained large language models (LLMs) to serve as dual-purpose cognitive models--capable of both accurate prediction and interpretable explanation in natural language. Specifically, we employ reinforcement learning with outcome-based rewards to guide LLMs toward generating explicit reasoning traces for explaining human risky choices. Our findings demonstrate that this approach produces high-quality explanations alongside strong quantitative predictions of human decisions.

📄 PDF Abstract BibTeX arXiv:2505.11614

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Few-shot Vision-based Human Activity Recognition with MLLM-based Visual Reinforcement Learning

2025-08-14 · Wenqi Zheng, Yutaka Arakawa arxiv

Reinforcement learning in large reasoning models enables learning from feedback on their outputs, making it particularly valuable in scenarios where fine-tuning data is limited. However, its application in multi-modal hu…

Human Activity RecognitionReinforcement Learning

A Survey on Explainable Deep Reinforcement Learning

2025-02-08 · Zelei Cheng, Jiahao Yu, Xinyu Xing

Deep Reinforcement Learning (DRL) has achieved remarkable success in sequential decision-making tasks across diverse domains, yet its reliance on black-box neural architectures hinders interpretability, trust, and deploy…

Adversarial RobustnessDecision MakingDeep Reinforcement Learningreinforcement-learning+3

SageLM: A Multi-aspect and Explainable Large Language Model for Speech Judgement

2025-08-28 · Yuan Ge, Junxiang Zhang, Xiaoqian Liu, Bei Li 외 arxiv

Speech-to-Speech (S2S) Large Language Models (LLMs) are foundational to natural human-computer interaction, enabling end-to-end spoken dialogue systems. However, evaluating these models remains a fundamental challenge. W…

Reinforcement Learning

TabReason: A Reinforcement Learning-Enhanced Reasoning LLM for Explainable Tabular Data Prediction

2025-05-27 · Tommy Xu, Zhitian Zhang, Xiangyu Sun, Lauren Kelly Zung 외

Predictive modeling on tabular data is the cornerstone of many real-world applications. Although gradient boosting machines and some recent deep models achieve strong performance on tabular data, they often lack interpre…

Learning from Explanations and Demonstrations: A Pilot Study

2020-11-01 · ACL (NL4XAI, INLG) 2020 11 · Silvia Tulli, Sebastian Wallkötter, Ana Paiva, Francisco S. Melo 외

AI has become prominent in a growing number of systems, and, as a direct consequence, the desire for explainability in such systems has become prominent as well. To build explainable systems, a large portion of existing …

text-to-speechText to SpeechTransfer Learning