paper-with-me

Papers

Principled Frameworks for Evaluating Ethics in NLP Systems

2019-06-14 · WS 2019 8 · Shrimai Prabhumoye, Elijah Mayfield, Alan W. black

We critique recent work on ethics in natural language processing. Those discussions have focused on data collection, experimental design, and interventions in modeling. But we argue that we ought to first understand the frameworks of ethics that are being used to evaluate the fairness and justice of algorithmic systems. Here, we begin that discussion by outlining deontological ethics, and envision a research agenda prioritized by it.

📄 PDF Abstract BibTeX arXiv:1906.06425

Code (0)

등록된 구현이 없습니다.

Tasks

EthicsExperimental DesignFairness

Similar Papers 제목 키워드 기반

ProMoral-Bench: Evaluating Prompting Strategies for Moral Reasoning and Safety in LLMs

2026-02-05 · Rohan Subramanian Thomas, Shikhar Shiromani, Abdullah Chaudhry, Ruizhe Li 외 arxiv

Prompt design significantly impacts the moral competence and safety alignment of large language models (LLMs), yet empirical comparisons remain fragmented across datasets and models.We introduce ProMoral-Bench, a unified…

Prompt Engineering

PsychEthicsBench: Evaluating Large Language Models Against Australian Mental Health Ethics

2026-01-07 · Yaling Shen, Stephanie Fong, Yiwen Jiang, Zimu Wang 외 arxiv

The increasing integration of large language models (LLMs) into mental health applications necessitates robust frameworks for evaluating professional safety alignment. Current evaluative approaches primarily rely on refu…

Reinforcement Learning and Machine ethics:a systematic review

2024-07-02 · Ajay Vishwanath, Louise A. Dennis, Marija Slavkovik

Machine ethics is the field that studies how ethical behaviour can be accomplished by autonomous systems. While there exist some systematic reviews aiming to consolidate the state of the art in machine ethics prior to 20…

Ethicsreinforcement-learningReinforcement Learning

Position: Stop Evaluating AI with Human Tests, Develop Principled, AI-specific Tests instead

2025-07-30 · Tom Sühr, Florian E. Dorner, Olawale Salaudeen, Augustin Kelava 외 arxiv

Large Language Models (LLMs) have achieved remarkable results on a range of standardized tests originally designed to assess human cognitive and psychological traits, such as intelligence and personality. While these res…

On the Efficiency of Ethics as a Governing Tool for Artificial Intelligence

2022-10-27 · Nicholas Kluge Corrêa, Nythamar de Oliveira, Diogo Massmann

The 4th Industrial Revolution is the culmination of the digital age. Nowadays, technologies such as robotics, nanotechnology, genetics, and artificial intelligence promise to transform our world and the way we live. Arti…

Ethics