paper-with-me

홈 › Papers

‘Am I the Bad One’? Predicting the Moral Judgement of the Crowd Using Pre–trained Language Models

2022-06-01 · LREC 2022 6 · Areej Alhassan, Jinkai Zhang, Viktor Schlegel

Natural language processing (NLP) has been shown to perform well in various tasks, such as answering questions, ascertaining natural language inference and anomaly detection. However, there are few NLP-related studies that touch upon the moral context conveyed in text. This paper studies whether state-of-the-art, pre-trained language models are capable of passing moral judgments on posts retrieved from a popular Reddit user board. Reddit is a social discussion website and forum where posts are promoted by users through a voting system. In this work, we construct a dataset that can be used for moral judgement tasks by collecting data from the AITA? (Am I the A*******?) subreddit. To model our task, we harnessed the power of pre-trained language models, including BERT, RoBERTa, RoBERTa-large, ALBERT and Longformer. We then fine-tuned these models and evaluated their ability to predict the correct verdict as judged by users for each post in the datasets. RoBERTa showed relative improvements across the three datasets, exhibiting a rate of 87% accuracy and a Matthews correlation coefficient (MCC) of 0.76, while the use of the Longformer model slightly improved the performance when used with longer sequences, achieving 87% accuracy and 0.77 MCC.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Anomaly DetectionNatural Language Inference

Similar Papers 제목 키워드 기반

Learning Moral Diversity: Modelling Individual Perspectives in Moral Classification of Texts

2026-06-22 · Yi Ren, Lewis Mitchell, Matthew Roughan arxiv

Understanding moral values in social media text offers insight into moral judgement formation, and supervised NLP models trained on crowdsourced data have achieved strong classification performance. However, most approac…

Decoding moral judgement from text: a pilot study

2024-05-28 · Diana E. Gherman, Thorsten O. Zander

Moral judgement is a complex human reaction that engages cognitive and emotional dimensions. While some of the morality neural correlates are known, it is currently unclear if we can detect moral violation at a single-tr…

Attribute

Are Language Models Sensitive to Morally Irrelevant Distractors?

2026-02-10 · Andrew Shaw, Christina Hahn, Catherine Rasgaitis, Yash Mishra 외 arxiv

With the rapid uptake of large language models (LLMs) across high-stakes settings, it is becoming increasingly important to ensure that LLMs behave in ways that align with human values. Existing moral benchmarks for this…

Moral Scenarios

Explainable Patterns for Distinction and Prediction of Moral Judgement on Reddit

2022-01-26 · Ion Stagkos Efstathiadis, Guilherme Paulino-Passos, Francesca Toni

The forum r/AmITheAsshole in Reddit hosts discussion on moral issues based on concrete narratives presented by users. Existing analysis of the forum focuses on its comments, and does not make the underlying data publicly…

Aligning AI With Shared Human Values

2020-08-05 · Dan Hendrycks, Collin Burns, Steven Basart, Andrew Critch 외

We show how to assess a language model's knowledge of basic concepts of morality. We introduce the ETHICS dataset, a new benchmark that spans concepts in justice, well-being, duties, virtues, and commonsense morality. Mo…

Ethicsreinforcement-learningReinforcement Learning (RL)World Knowledge