paper-with-me

Papers

A Risk Taxonomy for Evaluating AI-Powered Psychotherapy Agents

2025-05-21 · Ian Steenstra, Timothy W. Bickmore

The proliferation of Large Language Models (LLMs) and Intelligent Virtual Agents acting as psychotherapists presents significant opportunities for expanding mental healthcare access. However, their deployment has also been linked to serious adverse outcomes, including user harm and suicide, facilitated by a lack of standardized evaluation methodologies capable of capturing the nuanced risks of therapeutic interaction. Current evaluation techniques lack the sensitivity to detect subtle changes in patient cognition and behavior during therapy sessions that may lead to subsequent decompensation. We introduce a novel risk taxonomy specifically designed for the systematic evaluation of conversational AI psychotherapists. Developed through an iterative process including review of the psychotherapy risk literature, qualitative interviews with clinical and legal experts, and alignment with established clinical criteria (e.g., DSM-5) and existing assessment tools (e.g., NEQ, UE-ATR), the taxonomy aims to provide a structured approach to identifying and assessing user/patient harms. We provide a high-level overview of this taxonomy, detailing its grounding, and discuss potential use cases. We discuss two use cases in detail: monitoring cognitive model-based risk factors during a counseling conversation to detect unsafe deviations, in both human-AI counseling sessions and in automated benchmarking of AI psychotherapists with simulated patients. The proposed taxonomy offers a foundational step towards establishing safer and more responsible innovation in the domain of AI-driven mental health support.

📄 PDF Abstract BibTeX arXiv:2505.15108

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingDecompensation

Similar Papers 제목 키워드 기반

When LLM Therapists Become Salespeople: Evaluating Large Language Models for Ethical Motivational Interviewing

2025-03-30 · Haein Kong, Seonghyeon Moon

Large language models (LLMs) have been actively applied in the mental health field. Recent research shows the promise of LLMs in applying psychotherapy, especially motivational interviewing (MI). However, there is a lack…

EthicsResponse Generation

The Evolution of Alpha in Finance Harnessing Human Insight and LLM Agents

2025-05-20 · Mohammad Rubyet Islam

The pursuit of alpha returns that exceed market benchmarks has undergone a profound transformation, evolving from intuition-driven investing to autonomous, AI powered systems. This paper introduces a comprehensive five s…

Decision MakingRepresentation Learning

Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming

2026-02-23 · Ian Steenstra, Paola Pedrelli, Weiyan Shi, Stacy Marsella 외 arxiv

Large Language Models (LLMs) are increasingly utilized for mental health support; however, current safety benchmarks often fail to detect the complex, longitudinal risks inherent in therapeutic dialogue. We introduce an …

Red Teaming

A Survey of Large Language Models in Psychotherapy: Current Landscape and Future Directions

2025-02-16 · Hongbin Na, Yining Hua, Zimu Wang, Tao Shen 외

Mental health remains a critical global challenge, with increasing demand for accessible, effective interventions. Large language models (LLMs) offer promising solutions in psychotherapy by enhancing the assessment, diag…

Survey

Psychotherapy is Not One Thing: Simultaneous Modeling of Different Therapeutic Approaches

2022-07-01 · NAACL (CLPsych) 2022 7 · Maitrey Mehta, Derek Caperton, Katherine Axford, Lauren Weitzman 외

There are many different forms of psychotherapy. Itemized inventories of psychotherapeutic interventions provide a mechanism for evaluating the quality of care received by clients and for conducting research on how psych…

Multi-Label ClassificationMUlTI-LABEL-ClASSIFICATION