paper-with-me

Papers

Explainable Ethical Assessment on Human Behaviors by Generating Conflicting Social Norms

2025-12-16 · Yuxi Sun, Wei Gao, Hongzhan Lin, Jing Ma, Wenxuan Zhang arxiv

Human behaviors are often guided or constrained by social norms, which are defined as shared, commonsense rules. For example, underlying an action `\textit{report a witnessed crime}" are social norms that inform our conduct, such as \textit{It is expected to be brave to report crimes}''. Current AI systems that assess valence (i.e., support or oppose) of human actions by leveraging large-scale data training not grounded on explicit norms may be difficult to explain, and thus untrustworthy. Emulating human assessors by considering social norms can help AI models better understand and predict valence. While multiple norms come into play, conflicting norms can create tension and directly influence human behavior. For example, when deciding whether to `\textit{report a witnessed crime}'', one may balance \textit{bravery} against \textit{self-protection}. In this paper, we introduce \textit{ClarityEthic}, a novel ethical assessment approach, to enhance valence prediction and explanation by generating conflicting social norms behind human actions, which strengthens the moral reasoning capabilities of language models by using a contrastive learning strategy. Extensive experiments demonstrate that our method outperforms strong baseline approaches, and human evaluations confirm that the generated social norms provide plausible explanations for the assessment of human behaviors.

📄 PDF Abstract BibTeX arXiv:2512.15793

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive Learning

Similar Papers 제목 키워드 기반

ClarityEthic: Explainable Moral Judgment Utilizing Contrastive Ethical Insights from Large Language Models

2024-12-17 · Yuxi Sun, Wei Gao, Jing Ma, Hongzhan Lin 외

With the rise and widespread use of Large Language Models (LLMs), ensuring their safety is crucial to prevent harm to humans and promote ethical behaviors. However, directly assessing value valence (i.e., support or oppo…

Contrastive Learning

Explainable and Human-Grounded AI for Decision Support Systems: The Theory of Epistemic Quasi-Partnerships

2024-09-23 · John Dorsch, Maximilian Moll

In the context of AI decision support systems (AI-DSS), we argue that meeting the demands of ethical and explainable AI (XAI) is about developing AI-DSS to provide human decision-makers with three types of human-grounded…

PSYCHE: A Multi-faceted Patient Simulation Framework for Evaluation of Psychiatric Assessment Conversational Agents

2025-01-03 · Jingoo Lee, Kyungho Lim, Young-Chul Jung, Byung-Hoon Kim

Recent advances in large language models (LLMs) have accelerated the development of conversational agents capable of generating human-like responses. Since psychiatric assessments typically involve complex conversational…

Benchmarking

Democratizing Ethical Assessment of Natural Language Generation Models

2022-06-30 · Amin Rasekh, Ian Eisenberg

Natural language generation models are computer systems that generate coherent language when prompted with a sequence of words as context. Despite their ubiquity and many beneficial applications, language generation mode…

Text Generation

AI Act and Large Language Models (LLMs): When critical issues and privacy impact require human and ethical oversight

2024-03-31 · Nicola Fabiano

The imposing evolution of artificial intelligence systems and, specifically, of Large Language Models (LLM) makes it necessary to carry out assessments of their level of risk and the impact they may have in the area of p…