paper-with-me

Papers

DAIC-WOZ: On the Validity of Using the Therapist's prompts in Automatic Depression Detection from Clinical Interviews

2024-04-22 · Sergio Burdisso, Ernesto Reyes-Ramírez, Esaú Villatoro-Tello, Fernando Sánchez-Vega, Pastor López-Monroy, Petr Motlicek

Automatic depression detection from conversational data has gained significant interest in recent years. The DAIC-WOZ dataset, interviews conducted by a human-controlled virtual agent, has been widely used for this task. Recent studies have reported enhanced performance when incorporating interviewer's prompts into the model. In this work, we hypothesize that this improvement might be mainly due to a bias present in these prompts, rather than the proposed architectures and methods. Through ablation experiments and qualitative analysis, we discover that models using interviewer's prompts learn to focus on a specific region of the interviews, where questions about past experiences with mental health issues are asked, and use them as discriminative shortcuts to detect depressed participants. In contrast, models using participant responses gather evidence from across the entire interview. Finally, to highlight the magnitude of this bias, we achieve a 0.90 F1 score by intentionally exploiting it, the highest result reported to date on this dataset using only textual information. Our findings underline the need for caution when incorporating interviewers' prompts into models, as they may inadvertently learn to exploit targeted prompts, rather than learning to characterize the language and behavior that are genuinely indicative of the patient's mental health condition.

📄 PDF Abstract BibTeX arXiv:2404.14463

Code (1)

idiap/bias_in_daic-woz 공식 구현 pytorch

Tasks

Depression Detection

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

When Consistency Becomes Bias: Interviewer Effects in Semi-Structured Clinical Interviews

2026-03-25 · Hasindri Watawana, Sergio Burdisso, Diego A. Moreno-Galván, Fernando Sánchez-Vega 외 arxiv

Automatic depression detection from doctor-patient conversations has gained momentum thanks to the availability of public corpora and advances in language modeling. However, interpretability remains limited: strong perfo…

Gender Bias in Depression Detection Using Audio Features

2020-10-28 · Andrew Bailey, Mark D. Plumbley

Depression is a large-scale mental health problem and a challenging area for machine learning researchers in detection of depression. Datasets such as Distress Analysis Interview Corpus - Wizard of Oz (DAIC-WOZ) have bee…

BIG-bench Machine LearningDepression Detection

A Step Towards Preserving Speakers' Identity While Detecting Depression Via Speaker Disentanglement

2022-06-20 · Vijay Ravi, Jinhan Wang, Jonathan Flint, Abeer Alwan

Preserving a patient's identity is a challenge for automatic, speech-based diagnosis of mental health disorders. In this paper, we address this issue by proposing adversarial disentanglement of depression characteristics…

Depression DetectionDisentanglement

Confidence Estimation for Automatic Detection of Depression and Alzheimer's Disease Based on Clinical Interviews

2024-07-29 · Wen Wu, Chao Zhang, Philip C. Woodland

Speech-based automatic detection of Alzheimer's disease (AD) and depression has attracted increased attention. Confidence estimation is crucial for a trust-worthy automatic diagnostic system which informs the clinician a…

Diagnostic

Probabilistic Textual Time Series Depression Detection

2025-11-06 · Fabian Schmidt, Seyedehmoniba Ravan, Vladimir Vlassov arxiv

Accurate and interpretable predictions of depression severity are essential for clinical decision support, yet existing models often lack uncertainty estimates and temporal interpretability. We propose PTTSD, a Probabili…