paper-with-me

홈 › Papers

MentraSuite: Post-Training Large Language Models for Mental Health Reasoning and Assessment

2025-12-10 · Mengxi Xiao, Kailai Yang, Pengde Zhao, Enze Zhang, Ziyan Kuang, Zhiwei Liu, Weiguang Han, Shu Liao, Lianting Huang, Jinpeng Hu, Min Peng, Qianqian Xie, Sophia Ananiadou arxiv

Mental health disorders affect hundreds of millions globally, and the Web now serves as a primary medium for accessing support, information, and assessment. Large language models (LLMs) offer scalable and accessible assistance, yet their deployment in mental-health settings remains risky when their reasoning is incomplete, inconsistent, or ungrounded. Existing psychological LLMs emphasize emotional understanding or knowledge recall but overlook the step-wise, clinically aligned reasoning required for appraisal, diagnosis, intervention planning, abstraction, and verification. To address these issues, we introduce MentraSuite, a unified framework for advancing reliable mental-health reasoning. We propose MentraBench, a comprehensive benchmark spanning five core reasoning aspects, six tasks, and 13 datasets, evaluating both task performance and reasoning quality across five dimensions: conciseness, coherence, hallucination avoidance, task understanding, and internal consistency. We further present Mindora, a post-trained model optimized through a hybrid SFT-RL framework with an inconsistency-detection reward to enforce faithful and coherent reasoning. To support training, we construct high-quality trajectories using a novel reasoning trajectory generation strategy, that strategically filters difficult samples and applies a structured, consistency-oriented rewriting process to produce concise, readable, and well-balanced trajectories. Across 20 evaluated LLMs, Mindora achieves the highest average performance on MentraBench and shows remarkable performances in reasoning reliability, demonstrating its effectiveness for complex mental-health scenarios.

📄 PDF Abstract BibTeX arXiv:2512.09636

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Continual Training of Language Models for Few-Shot Learning

2022-10-11 · Zixuan Ke, Haowei Lin, Yijia Shao, Hu Xu 외

Recent work on applying large language models (LMs) achieves impressive performance in many NLP applications. Adapting or posttraining an LM using an unlabeled domain corpus can produce even better performance for end-ta…

Continual LearningContinual PretrainingFew-Shot LearningLanguage Modelling

Post-training for Deepfake Speech Detection

2025-06-26 · Wanying Ge, Xin Wang, Xuechen Liu, Junichi Yamagishi

We introduce a post-training approach that adapts self-supervised learning (SSL) models for deepfake speech detection by bridging the gap between general pre-training and domain-specific fine-tuning. We present AntiDeepf…

Face SwappingSelf-Supervised Learning

Understanding LLMs' Cross-Lingual Context Retrieval: How Good It Is And Where It Comes From

2025-04-15 · Changjiang Gao, Hankun Lin, ShuJian Huang, Xin Huang 외

The ability of cross-lingual context retrieval is a fundamental aspect of cross-lingual alignment of large language models (LLMs), where the model extracts context information in one language based on requests in another…

Machine Reading ComprehensionReading ComprehensionRetrieval

Saten: Sparse Augmented Tensor Networks for Post-Training Compression of Large Language Models

2025-05-20 · Ryan Solgi, Kai Zhen, Rupak Vignesh Swaminathan, Nathan Susanj 외

The efficient implementation of large language models (LLMs) is crucial for deployment on resource-constrained devices. Low-rank tensor compression techniques, such as tensor-train (TT) networks, have been widely studied…

Model CompressionTensor Networks

Understanding Post-Training Structural Changes in Large Language Models

2025-09-22 · Xinyu He, Xianghui Cao arxiv

Post-training fundamentally alters the behavior of large language models (LLMs), yet its impact on the internal parameter space remains poorly understood. In this work, we conduct a systematic singular value decompositio…