paper-with-me

Papers

Towards interactive evaluations for interaction harms in human-AI systems

2024-05-17 · Lujain Ibrahim, Saffron Huang, Umang Bhatt, Lama Ahmad, Markus Anderljung

Current AI evaluation paradigms that rely on static, model-only tests fail to capture harms that emerge through sustained human-AI interaction. As interactive AI systems, such as AI companions, proliferate in daily life, this mismatch between evaluation methods and real-world use becomes increasingly consequential. We argue for a paradigm shift toward evaluation centered on \textit{interactional ethics}, which addresses risks like inappropriate human-AI relationships, social manipulation, and cognitive overreliance that develop through repeated interaction rather than single outputs. Drawing on human-computer interaction, natural language processing, and the social sciences, we propose principles for evaluating generative models through interaction scenarios and human impact metrics. We conclude by examining implementation challenges and open research questions for researchers, practitioners, and regulators integrating these approaches into AI governance frameworks.

📄 PDF Abstract BibTeX arXiv:2405.10632

Code (0)

등록된 구현이 없습니다.

Tasks

Ethics

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Interactive Counterfactual Exploration of Algorithmic Harms in Recommender Systems

2024-09-10 · Yongsu Ahn, Quinn K Wolter, Jonilyn Dick, Janet Dick 외

Recommender systems have become integral to digital experiences, shaping user interactions and preferences across various platforms. Despite their widespread use, these systems often suffer from algorithmic biases that c…

counterfactualFairnessRecommendation Systems

FairPrism: Evaluating Fairness-Related Harms in Text Generation

2023-07-01 · Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics 2023 7 · Eve Fleisig, Aubrie Amstutz, Chad Atalla, Su Lin Blodgett 외

It is critical to measure and mitigate fairness- related harms caused by AI text generation systems, including stereotyping and demeaning harms. To that end, we introduce FairPrism, a dataset of 5,000 examples of AI-gene…

FairnessText Generation

Measuring What Matters: Connecting AI Ethics Evaluations to System Attributes, Hazards, and Harms

2025-10-11 · Shalaleh Rismani, Renee Shelby, Leah Davis, Negar Rostamzadeh 외 arxiv

Over the past decade, an ecosystem of measures has emerged to evaluate the social and ethical implications of AI systems, largely shaped by high-level ethics principles. These measures are developed and used in fragmente…

Not My Voice! A Taxonomy of Ethical and Safety Harms of Speech Generators

2024-01-25 · Wiebke Hutiri, Oresiti Papakyriakopoulos, Alice Xiang

The rapid and wide-scale adoption of AI to generate human speech poses a range of significant ethical and safety risks to society that need to be addressed. For example, a growing number of speech generation incidents ar…

Decision Making

MDrive: Benchmarking Closed-Loop Cooperative Driving for End-to-End Multi-agent Systems

2026-05-11 · Marco Coscoy, Zewei Zhou, Seth Z. Zhao, Henry Wei 외 arxiv

Vehicle-to-Everything (V2X) communication has emerged as a promising paradigm for autonomous driving, enabling connected agents to share complementary perception information and negotiate with each other to benefit the f…

Autonomous Driving