Language Models are Alignable Decision-Makers: Dataset and Application to the Medical Triage Domain
In difficult decision-making scenarios, it is common to have conflicting opinions among expert human decision-makers as there may not be a single right answer. Such decisions may be guided by different attributes that can be used to characterize an individual's decision. We introduce a novel dataset for medical triage decision-making, labeled with a set of decision-maker attributes (DMAs). This dataset consists of 62 scenarios, covering six different DMAs, including ethical principles such as fairness and moral desert. We present a novel software framework for human-aligned decision-making by utilizing these DMAs, paving the way for trustworthy AI with better guardrails. Specifically, we demonstrate how large language models (LLMs) can serve as ethical decision-makers, and how their decisions can be aligned to different DMAs using zero-shot prompting. Our experiments focus on different open-source models with varying sizes and training techniques, such as Falcon, Mistral, and Llama 2. Finally, we also introduce a new form of weighted self-consistency that improves the overall quantified performance. Our results provide new research directions in the use of LLMs as alignable decision-makers. The dataset and open-source software are publicly available at: https://github.com/ITM-Kitware/llm-alignable-dm.
Code (1)
Tasks
Decision MakingFairnessMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Sync from the Sea: Retrieving Alignable Videos from Large-Scale Datasets
Temporal video alignment aims to synchronize the key events like object interactions or action phase transitions in two videos. Such methods could benefit various video editing, processing, and understanding tasks. Howev…
Video AlignmentVideo EditingVideo RetrievalGRIFFIN: Effective Token Alignment for Faster Speculative Decoding
Speculative decoding accelerates inference in large language models (LLMs) by generating multiple draft tokens simultaneously. However, existing methods often struggle with token misalignment between the training and dec…
Analysis and Prediction of Unalignable Words in Parallel Text
(Un)certainty of (Un)fairness: Preference-Based Selection of Certainly Fair Decision-Makers
Fairness metrics are used to assess discrimination and bias in decision-making processes across various domains, including machine learning models and human decision-makers in real-world applications. This involves calcu…
Decision MakingFairnessAn Information-theoretic On-line Learning Principle for Specialization in Hierarchical Decision-Making Systems
Information-theoretic bounded rationality describes utility-optimizing decision-makers whose limited information-processing capabilities are formalized by information constraints. One of the consequences of bounded ratio…
Decision MakingReinforcement Learning