paper-with-me

홈 › Papers

Domain-Specific Constitutional AI: Enhancing Safety in LLM-Powered Mental Health Chatbots

2025-09-19 · Chenhan Lyu, Yutong Song, Pengfei Zhang, Amir M. Rahmani arxiv

Mental health applications have emerged as a critical area in computational health, driven by rising global rates of mental illness, the integration of AI in psychological care, and the need for scalable solutions in underserved communities. These include therapy chatbots, crisis detection, and wellness platforms handling sensitive data, requiring specialized AI safety beyond general safeguards due to emotional vulnerability, risks like misdiagnosis or symptom exacerbation, and precise management of vulnerable states to avoid severe outcomes such as self-harm or loss of trust. Despite AI safety advances, general safeguards inadequately address mental health-specific challenges, including crisis intervention accuracy to avert escalations, therapeutic guideline adherence to prevent misinformation, scale limitations in resource-constrained settings, and adaptation to nuanced dialogues where generics may introduce biases or miss distress signals. We introduce an approach to apply Constitutional AI training with domain-specific mental health principles for safe, domain-adapted CAI systems in computational mental health applications.

📄 PDF Abstract BibTeX arXiv:2509.16444

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Glass Box at Orbit: A Constitutional AI Verification Framework for Trustworthy Autonomous CubeSat Intelligence

2026-06-02 · Karthik Barma, Anil Sanneboyina, V C Premchand Yadav arxiv

The space industry is quietly building toward something nobody has fully reckoned with: orbital data centers running thousands of autonomous AI workloads with no human in the loop, 550 km above the Earth. Microsoft, AWS,…

Constitutional On-Policy Safe Distillation

2026-06-02 · Ming Wen, Yuxuan Liu, Kun Yang, Yunhao Feng 외 arxiv

On-policy self-distillation (OPSD) has emerged as an efficient post-training paradigm by using a teacher conditioned on privileged information to provide dense token-level supervision. Prior work has shown that OPSD can …

Toward Responsible Federated Large Language Models: Leveraging a Safety Filter and Constitutional AI

2025-02-23 · Eunchung Noh, Jeonghun Baek

Recent research has increasingly focused on training large language models (LLMs) using federated learning, known as FedLLM. However, responsible AI (RAI), which aims to ensure safe responses, remains underexplored in th…

Federated Learning

Constitutional Spec-Driven Development: Enforcing Security by Construction in AI-Assisted Code Generation

2026-01-31 · Srinivas Rao Marri arxiv

The proliferation of AI-assisted "vibe coding" enables rapid software development but introduces significant security risks, as Large Language Models (LLMs) prioritize functional correctness over security. We present Con…

Code Generation

Enhancing Robustness of LLM-Driven Multi-Agent Systems through Randomized Smoothing

2025-07-05 · Jinwei Hu, Yi Dong, Zhengtao Ding, Xiaowei Huang arxiv

This paper presents a defense framework for enhancing the safety of large language model (LLM) empowered multi-agent systems (MAS) in safety-critical domains such as aerospace. We apply randomized smoothing, a statistica…

Computational Efficiency