paper-with-me

홈 › Papers

SENSE-7: Taxonomy and Dataset for Measuring User Perceptions of Empathy in Sustained Human-AI Conversations

2025-09-19 · Jina Suh, Lindy Le, Erfan Shayegani, Gonzalo Ramos, Judith Amores, Desmond C. Ong, Mary Czerwinski, Javier Hernandez arxiv

Empathy is increasingly recognized as a key factor in human-AI communication, yet conventional approaches to "digital empathy" often focus on simulating internal, human-like emotional states while overlooking the inherently subjective, contextual, and relational facets of empathy as perceived by users. In this work, we propose a human-centered taxonomy that emphasizes observable empathic behaviors and introduce a new dataset, Sense-7, of real-world conversations between information workers and Large Language Models (LLMs), which includes per-turn empathy annotations directly from the users, along with user characteristics, and contextual details, offering a more user-grounded representation of empathy. Analysis of 695 conversations from 109 participants reveals that empathy judgments are highly individualized, context-sensitive, and vulnerable to disruption when conversational continuity fails or user expectations go unmet. To promote further research, we provide a subset of 672 anonymized conversation and provide exploratory classification analysis, showing that an LLM-based classifier can recognize 5 levels of empathy with an encouraging average Spearman $ρ$=0.369 and Accuracy=0.487 over this set. Overall, our findings underscore the need for AI designs that dynamically tailor empathic behaviors to user contexts and goals, offering a roadmap for future research and practical development of socially attuned, human-centered artificial agents.

📄 PDF Abstract BibTeX arXiv:2509.16437

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

HumT DumT: Measuring and controlling human-like language in LLMs

2025-02-18 · Myra Cheng, Sunny Yu, Dan Jurafsky

Should LLMs generate language that makes them seem human? Human-like language might improve user experience, but might also lead to deception, overreliance, and stereotyping. Assessing these potential impacts requires a …

Text Generation

User Perceptions vs. Proxy LLM Judges: Privacy and Helpfulness in LLM Responses to Privacy-Sensitive Scenarios

2025-10-23 · Xiaoyuan Wu, Roshni Kaushik, Wenkai Li, Lujo Bauer 외 arxiv

Large language models (LLMs) are rapidly being adopted for tasks like drafting emails, summarizing meetings, and answering health questions. In these settings, users may need to share private information (e.g., contact d…

Decoy Effect In Search Interaction: Understanding User Behavior and Measuring System Vulnerability

2024-03-27 · Nuo Chen, Jiqun Liu, Hanpei Fang, Yuankai Luo 외

This study examines the decoy effect's underexplored influence on user search interactions and methods for measuring information retrieval (IR) systems' vulnerability to this effect. It explores how decoy results alter u…

Information RetrievalRetrieval

Visualizing word senses in WordNet Atlas

2012-05-01 · LREC 2012 5 · Matteo Abrate, Clara Bacciu

This demo presents the second prototype of WordNet Atlas, a web application that gives users the ability to navigate and visualize the 146,312 word senses of the nouns contained within the Princeton WordNet. Two compleme…

LEMMANavigate

Ask LLMs Directly, "What shapes your bias?": Measuring Social Bias in Large Language Models

2024-06-06 · Jisu Shin, Hoyun Song, Huije Lee, Soyeong Jeong 외

Social bias is shaped by the accumulation of social perceptions towards targets across various demographic identities. To fully understand such social bias in large language models (LLMs), it is essential to consider the…