Histoires Morales: A French Dataset for Assessing Moral Alignment
Aligning language models with human values is crucial, especially as they become more integrated into everyday life. While models are often adapted to user preferences, it is equally important to ensure they align with moral norms and behaviours in real-world social situations. Despite significant progress in languages like English and Chinese, French has seen little attention in this area, leaving a gap in understanding how LLMs handle moral reasoning in this language. To address this gap, we introduce Histoires Morales, a French dataset derived from Moral Stories, created through translation and subsequently refined with the assistance of native speakers to guarantee grammatical accuracy and adaptation to the French cultural context. We also rely on annotations of the moral values within the dataset to ensure their alignment with French norms. Histoires Morales covers a wide range of social situations, including differences in tipping practices, expressions of honesty in relationships, and responsibilities toward animals. To foster future research, we also conduct preliminary experiments on the alignment of multilingual models on French and English data and the robustness of the alignment. We find that while LLMs are generally aligned with human moral norms by default, they can be easily influenced with user-preference optimization for both moral and immoral data.
Code (1)
Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Some game theoretic marketing attribution models
In this paper, we propose and analyse two game theoretical models useful to design marketing channels attribution mechanisms based on cooperative TU games and bankruptcy problems, respectively. First, we analyse the Sum …
MarketingLocalized Questions in Medical Visual Question Answering
Visual Question Answering (VQA) models aim to answer natural language questions about given images. Due to its ability to ask questions that differ from those used when training the model, medical VQA has received substa…
Medical Visual Question AnsweringQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)Targeted Visual Prompting for Medical Visual Question Answering
With growing interest in recent years, medical visual question answering (Med-VQA) has rapidly evolved, with multimodal large language models (MLLMs) emerging as an alternative to classical model architectures. Specifica…
Medical Visual Question AnsweringQuestion AnsweringVisual PromptingVisual Question Answering+1Precluding Oscillations in Michaelis-Menten Approximations of Dual-site Phosphorylation Systems
Oscillations play a major role in a number of biological systems, from predator-prey models of ecology to circadian clocks. In this paper we focus on the question of whether oscillations exist within dual-site phosphoryl…
Evaluating Moral Beliefs across LLMs through a Pluralistic Framework
Proper moral beliefs are fundamental for language models, yet assessing these beliefs poses a significant challenge. This study introduces a novel three-module framework to evaluate the moral beliefs of four prominent la…