paper-with-me

홈 › Papers

Probing the "Psyche'' of Large Reasoning Models: Understanding Through a Human Lens

2025-11-30 · Yuxiang Chen, Zuohan Wu, Ziwei Wang, Xiangning Yu, Xujia Li, Linyi Yang, Mengyue Yang, Jun Wang, Lei Chen arxiv

Large reasoning models (LRMs) have garnered significant attention from researchers owing to their exceptional capability in addressing complex tasks. Motivated by the observed human-like behaviors in their reasoning processes, this paper introduces a comprehensive taxonomy to characterize atomic reasoning steps and probe the `psyche'' of LRM intelligence. Specifically, it comprises five groups and seventeen categories derived from human mental processes, thereby grounding the understanding of LRMs in an interdisciplinary perspective. The taxonomy is then applied for an in-depth understanding of current LRMs, resulting in a distinct labeled dataset that comprises 277,534 atomic reasoning steps. Using this resource, we analyze contemporary LRMs and distill several actionable takeaways for improving training and post-training of reasoning models. Notably, our analysis reveals that prevailing post-answer `double-checks'' (self-monitoring evaluations) are largely superficial and rarely yield substantive revisions. Thus, incentivizing comprehensive multi-step reflection, rather than simple self-monitoring, may offer a more effective path forward. To complement the taxonomy, an automatic annotation framework, named CAPO, is proposed to leverage large language models (LLMs) for generating the taxonomy-based annotations. Experimental results demonstrate that CAPO achieves higher consistency with human experts compared to baselines, facilitating a scalable and comprehensive analysis of LRMs from a human cognitive perspective. Together, the taxonomy, CAPO, and the derived insights provide a principled, scalable path toward understanding and advancing LRM reasoning.

📄 PDF Abstract BibTeX arXiv:2512.00729

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Psyche-R1: Towards Reliable Psychological LLMs through Unified Empathy, Expertise, and Reasoning

2025-08-14 · Chongyuan Dai, Jinpeng Hu, Hongchang Shi, Zhuo Li 외 arxiv

Amidst a shortage of qualified mental health professionals, the integration of large language models (LLMs) into psychological applications offers a promising way to alleviate the growing burden of mental health disorder…

Empathetic Response Generation

Neuroplasticity and Psychedelics: a comprehensive examination of classic and non-classic compounds in pre and clinical models

2024-11-29 · Claudio Agnorelli, Meg Spriggs, Kate Godfrey, Gabriela Sawicka 외

Neuroplasticity, the ability of the nervous system to adapt throughout an organism's lifespan, offers potential as both a biomarker and treatment target for neuropsychiatric conditions. Psychedelics, a burgeoning categor…

How to set up a psychedelic study: Unique considerations for research involving human participants

2025-03-28 · Marcus J. Glennon, Catherine I. V. Bird, Prateek Yadav, Patrick Kleine 외

Setting up a psychedelic study can be a long, arduous, and kafkaesque process. This rapidly-developing field poses several unique challenges for researchers, necessitating a range of considerations that have not yet been…

PSYCHE: A Multi-faceted Patient Simulation Framework for Evaluation of Psychiatric Assessment Conversational Agents

2025-01-03 · Jingoo Lee, Kyungho Lim, Young-Chul Jung, Byung-Hoon Kim

Recent advances in large language models (LLMs) have accelerated the development of conversational agents capable of generating human-like responses. Since psychiatric assessments typically involve complex conversational…

Benchmarking

PsychEthicsBench: Evaluating Large Language Models Against Australian Mental Health Ethics

2026-01-07 · Yaling Shen, Stephanie Fong, Yiwen Jiang, Zimu Wang 외 arxiv

The increasing integration of large language models (LLMs) into mental health applications necessitates robust frameworks for evaluating professional safety alignment. Current evaluative approaches primarily rely on refu…