paper-with-me

홈 › Papers

MOSAIC: Unveiling the Moral, Social and Individual Dimensions of Large Language Models

2026-02-09 · Erica Coppolillo, Emilio Ferrara arxiv

Large Language Models (LLMs) are increasingly deployed in sensitive applications including psychological support, healthcare, and high-stakes decision-making. This expansion has motivated growing research into the ethical and moral foundations underlying LLM behavior, raising critical questions about their reliability in ethical reasoning. However, existing studies and benchmarks rely almost exclusively on Moral Foundation Theory (MFT), largely neglecting other relevant dimensions such as social values, personality traits, and individual characteristics that shape human ethical reasoning. To address these limitations, we introduce MOSAIC, the first large-scale benchmark designed to jointly assess the moral, social, and individual characteristics of LLMs. The benchmark comprises nine validated questionnaires drawn from moral philosophy, psychology, and social theory, alongside four platform-based games designed to probe morally ambiguous scenarios. In total, MOSAIC includes over 600 curated questions and scenarios, released as a ready-to-use, extensible resource for evaluating the behavioral foundations of LLMs. We validate the benchmark across three models from different families, demonstrating its utility across all assessed dimensions and providing the first empirical evidence that MFT alone is insufficient to comprehensively evaluate complex AI systems' ethical behavior. We publicly release the dataset and our benchmark Python library.

📄 PDF Abstract BibTeX arXiv:2603.00048

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Enhancing Stance Classification on Social Media Using Quantified Moral Foundations

2023-10-15 · Hong Zhang, Quoc-Nam Nguyen, Prasanta Bhattacharya, Wei Gao 외

This study enhances stance detection on social media by incorporating deeper psychological attributes, specifically individuals' moral foundations. These theoretically-derived dimensions aim to provide a comprehensive pr…

ClassificationStance ClassificationStance Detection

A Computational Model of Commonsense Moral Decision Making

2018-01-12 · Richard Kim, Max Kleiman-Weiner, Andres Abeliuk, Edmond Awad 외

We introduce a new computational model of moral decision making, drawing on a recent theory of commonsense moral learning via social dynamics. Our model describes moral dilemmas as a utility function that computes trade-…

Autonomous VehiclesDecision Makingmodel

Morality-based Assertion and Homophily on Social Media: A Cultural Comparison between English and Japanese Languages

2021-08-24 · Maneet Singh, Rishemjit Kaur, Akiko Matsuo, S. R. S. Iyengar 외

Moral psychology is a domain that deals with moral identity, appraisals and emotions. Previous work has primarily focused on moral development and the associated role of culture. Knowing that language is an inherent elem…

Cultural Vocal Bursts Intensity PredictionFairness

A Survey on Moral Foundation Theory and Pre-Trained Language Models: Current Advances and Challenges

2024-09-20 · Lorenzo Zangari, Candida M. Greco, Davide Picca, Andrea Tagarelli

Moral values have deep roots in early civilizations, codified within norms and laws that regulated societal order and the common good. They play a crucial role in understanding the psychological basis of human behavior a…

Evolution of social norms for moral judgment

2022-04-22 · Taylor A. Kessinger, Corina E. Tarnita, Joshua B. Plotkin

Reputations provide a powerful mechanism to sustain cooperation, as individuals cooperate with those of good social standing. But how should moral reputations be updated as we observe social behavior, and when will a pop…