paper-with-me

Papers

When Consistency Becomes Bias: Interviewer Effects in Semi-Structured Clinical Interviews

2026-03-25 · Hasindri Watawana, Sergio Burdisso, Diego A. Moreno-Galván, Fernando Sánchez-Vega, A. Pastor López-Monroy, Petr Motlicek, Esaú Villatoro-Tello arxiv

Automatic depression detection from doctor-patient conversations has gained momentum thanks to the availability of public corpora and advances in language modeling. However, interpretability remains limited: strong performance is often reported without revealing what drives predictions. We analyze three datasets: ANDROIDS, DAIC-WOZ, E-DAIC and identify a systematic bias from interviewer prompts in semi-structured interviews. Models trained on interviewer turns exploit fixed prompts and positions to distinguish depressed from control subjects, often achieving high classification scores without using participant language. Restricting models to participant utterances distributes decision evidence more broadly and reflects genuine linguistic cues. While semi-structured protocols ensure consistency, including interviewer prompts inflates performance by leveraging script artifacts. Our results highlight a cross-dataset, architecture-agnostic bias and emphasize the need for analyses that localize decision evidence by time and speaker to ensure models learn from participants' language.

📄 PDF Abstract BibTeX arXiv:2603.24651

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Exploring the Implementation of AI in Early Onset Interviews to Help Mitigate Bias

2025-01-17 · Nishka Lal, Omar Benkraouda

This paper investigates the application of artificial intelligence (AI) in early-stage recruitment interviews in order to reduce inherent bias, specifically sentiment bias. Traditional interviewers are often subject to s…

LLM-as-an-Interviewer: Beyond Static Testing Through Dynamic LLM Evaluation

2024-12-10 · Eunsu Kim, Juyoung Suk, Seungone Kim, Niklas Muennighoff 외

We introduce LLM-as-an-Interviewer, a novel paradigm for evaluating large language models (LLMs). This approach leverages multi-turn interactions where the LLM interviewer actively provides feedback on responses and pose…

Math

DAIC-WOZ: On the Validity of Using the Therapist's prompts in Automatic Depression Detection from Clinical Interviews

2024-04-22 · Sergio Burdisso, Ernesto Reyes-Ramírez, Esaú Villatoro-Tello, Fernando Sánchez-Vega 외

Automatic depression detection from conversational data has gained significant interest in recent years. The DAIC-WOZ dataset, interviews conducted by a human-controlled virtual agent, has been widely used for this task.…

Depression Detection

"I ain't tellin' white folks nuthin": A quantitative exploration of the race-related problem of candour in the WPA slave narratives

2018-05-01 · Soumya Kambhampati

From 1936-38, the Works Progress Administration interviewed thousands of former slaves about their life experiences. While these interviews are crucial to understanding the "peculiar institution" from the standpoint of t…

Sentiment Analysis

Mic Drop or Data Flop? Evaluating the Fitness for Purpose of AI Voice Interviewers for Data Collection within Quantitative & Qualitative Research Contexts

2025-09-01 · Shreyas Tirumala, Nishant Jain, Danny D. Leybzon, Trent D. Buskirk arxiv

Transformer-based Large Language Models (LLMs) have paved the way for "AI interviewers" that can administer voice-based surveys with respondents in real-time. This position paper reviews emerging evidence to understand w…

Speech Recognition