paper-with-me

홈 › Papers

The Moralization Corpus: Frame-Based Annotation and Analysis of Moralizing Speech Acts across Diverse Text Genres

2025-12-17 · Maria Becker, Mirko Sommer, Lars Tapken, Yi Wan Teh, Bruno Brocai arxiv

Moralizations - arguments that invoke moral values to justify demands or positions - are a yet underexplored form of persuasive communication. We present the Moralization Corpus, a novel multi-genre dataset designed to analyze how moral values are strategically used in argumentative discourse. Moralizations are pragmatically complex and often implicit, posing significant challenges for both human annotators and NLP systems. We develop a frame-based annotation scheme that captures the constitutive elements of moralizations - moral values, demands, and discourse protagonists - and apply it to a diverse set of German texts, including political debates, news articles, and online discussions. The corpus enables fine-grained analysis of moralizing language across communicative formats and domains. We further evaluate several large language models (LLMs) under varied prompting conditions for the task of moralization detection and moralization component extraction and compare it to human annotations in order to investigate the challenges of automatic and manual analysis of moralizations. Results show that detailed prompt instructions has a greater effect than few-shot or explanation-based prompting, and that moralization remains a highly subjective and context-sensitive task. We release all data, annotation guidelines, and code to foster future interdisciplinary research on moral discourse and moral reasoning in NLP.

📄 PDF Abstract BibTeX arXiv:2512.15248

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A data and analysis resource for an experiment in text mining a collection of micro-blogs on a political topic.

2012-05-01 · LREC 2012 5 · William Black, Rob Procter, Steven Gray, Sophia Ananiadou

The analysis of a corpus of micro-blogs on the topic of the 2011 UK referendum about the Alternative Vote has been undertaken as a joint activity by text miners and social scientists. To facilitate the collaboration, the…

Named Entity Recognition (NER)Sentiment Analysistext annotation

Improving corpus annotation productivity: a method and experiment with interactive tagging

2012-05-01 · LREC 2012 5 · Atro Voutilainen

Corpus linguistic and language technological research needs empirical corpus data with nearly correct annotation and high volume to enable advances in language modelling and theorising. Recent work on improving corpus an…

Language Modelling

MLSA --- A Multi-layered Reference Corpus for German Sentiment Analysis

2012-05-01 · LREC 2012 5 · Simon Clematide, Stefan Gindl, Manfred Klenner, Stefanos Petrakis 외

In this paper, we describe MLSA, a publicly available multi-layered reference corpus for German-language sentiment analysis. The construction of the corpus is based on the manual annotation of 270 German-language sentenc…

Opinion MiningQuestion AnsweringSentenceSentiment Analysis

Testing Focus and Non-at-issue Frameworks with a Question-under-Discussion-Annotated Corpus

2022-06-01 · LREC 2022 6 · Christoph Hesse, Maurice Langner, Ralf Klabunde, Anton Benz

We present an annotated corpus of German driving reports for the analysis of Question-under-Discussion (QUD) based information structural distinctions. Since QUDs can hardly be defined in advance for providing a correspo…

NarrativeTime: Dense Temporal Annotation on a Timeline

2019-08-29 · Anna Rogers, Marzena Karpinska, Ankita Gupta, Vladislav Lialin 외

For the past decade, temporal annotation has been sparse: only a small portion of event pairs in a text was annotated. We present NarrativeTime, the first timeline-based annotation framework that achieves full coverage o…

Chunking