paper-with-me

Papers

Responsible Evaluation of AI for Mental Health

2026-01-20 · Hiba Arnaout, Anmol Goel, H. Andrew Schwartz, Steffen T. Eberhardt, Dana Atzil-Slonim, Gavin Doherty, Brian Schwartz, Wolfgang Lutz, Tim Althoff, Munmun De Choudhury, Hamidreza Jamalabadi, Raj Sanjay Shah, Flor Miriam Plaza-del-Arco, Dirk Hovy, Maria Liakata, Iryna Gurevych arxiv

Although artificial intelligence (AI) shows growing promise for mental health care, current approaches to evaluating AI tools in this domain remain fragmented and poorly aligned with clinical practice, social context, and first-hand user experience. This paper argues for a rethinking of responsible evaluation -- what is measured, by whom, and for what purpose -- by introducing an interdisciplinary framework that integrates clinical soundness, social context, and equity, providing a structured basis for evaluation. Through an analysis of 135 recent *CL publications, we identify recurring limitations, including over-reliance on generic metrics that do not capture clinical validity, therapeutic appropriateness, or user experience, limited participation from mental health professionals, and insufficient attention to safety and equity. To address these gaps, we propose a taxonomy of AI mental health support types -- assessment-, intervention-, and information synthesis-oriented -- each with distinct risks and evaluative requirements, and illustrate its use through case studies.

📄 PDF Abstract BibTeX arXiv:2602.00065

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

VERA-MH: Validation of Ethical and Responsible AI in Mental Health

2026-05-13 · Luca Belli, Kate H. Bentley, Josh Gieringer, Emily Van Ark 외 arxiv

Chatbot usage has increased, including in fields for which they were never developed for--notably mental health support. To that end, we introduce Validations of Ethical and Responsible AI in Mental Health (VERA-MH), a n…

Position: Beyond Assistance -- Reimagining LLMs as Ethical and Adaptive Co-Creators in Mental Health Care

2025-02-21 · Abeer Badawi, Md Tahmid Rahman Laskar, Jimmy Xiangji Huang, Shaina Raza 외

This position paper argues for a fundamental shift in how Large Language Models (LLMs) are integrated into the mental health care domain. We advocate for their role as co-creators rather than mere assistive tools. While …

Position

RHealthTwin: Towards Responsible and Multimodal Digital Twins for Personalized Well-being

2025-06-10 · Rahatara Ferdousi, M Anwar Hossain

The rise of large language models (LLMs) has created new possibilities for digital twins in healthcare. However, the deployment of such systems in consumer health contexts raises significant concerns related to hallucina…

HallucinationInstruction FollowingNutrition

Mentalic Net: Development of RAG-based Conversational AI and Evaluation Framework for Mental Health Support

2025-08-27 · Anandi Dutta, Shivani Mruthyunjaya, Jessica Saddington, Kazi Sifatul Islam arxiv

The emergence of large language models (LLMs) has unlocked boundless possibilities, along with significant challenges. In response, we developed a mental health support chatbot designed to augment professional healthcare…

Prompt Engineering

RAISE: A Unified Framework for Responsible AI Scoring and Evaluation

2025-10-21 · Loc Phuc Truong Nguyen, Hung Thanh Do arxiv

As AI systems enter high-stakes domains, evaluation must extend beyond predictive accuracy to include explainability, fairness, robustness, and sustainability. We introduce RAISE (Responsible AI Scoring and Evaluation), …