paper-with-me

홈 › Papers

COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought Completion

2026-05-14 · Zihan Deng, Xiaozhen Zhong, Chuanzhi Xu arxiv

As large language models empower healthcare, intelligent clinical decision support has developed rapidly. Longitudinal electronic health records (EHR) provide essential temporal evidence for accurate clinical diagnosis and analysis. However, current large language models have critical flaws in longitudinal EHR reasoning. First, lacking fine-grained statistical reasoning, they often hallucinate clinical trends and metrics when quantitative evidence is textually implied, biasing diagnostic inference. Second, non-uniform time series and scarce labels in longitudinal EHR hinder models from capturing long-range temporal dependencies, limiting reliable clinical reasoning. To address the above limitations, this work presents the Probabilistic Chain-of-Thought Completion Agent (COTCAgent), a hierarchical reasoning framework for longitudinal electronic health records. It consists of three core modules. The Temporal-Statistics Adapter (TSA) converts analytical plans into executable code for standardized trend output. The Chain-of-Thought Completion (COTC) layer leverages a symptom-trend-disease knowledge base with weighted scoring to evaluate disease risk, while the bounded completion module acquires structured evidence through standardized inquiries and iterative scoring constraints to ensure rigorous reasoning. By decoupling statistical computation, feature matching, and language generation, the framework eliminates reliance on complex multi-modal inputs and enables efficient longitudinal record analysis with lower computational overhead. Experimental results show that COTCAgent powered by Baichuan-M2 achieves 90.47% Top-1 accuracy on the self-built dataset and 70.41% on HealthBench, outperforming existing medical agents and mainstream large language models. The code is available at https://github.com/FrankDengAI/COTCAgent/.

📄 PDF Abstract BibTeX arXiv:2605.15016

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MDTeamGPT: A Self-Evolving LLM-based Multi-Agent Framework for Multi-Disciplinary Team Medical Consultation

2025-03-18 · Kai Chen, Xinfeng Li, Tianpei Yang, Hewei Wang 외

Large Language Models (LLMs) have made significant progress in various fields. However, challenges remain in Multi-Disciplinary Team (MDT) medical consultations. Current research enhances reasoning through role assignmen…

MedQA

Psychologically-informed chain-of-thought prompts for metaphor understanding in large language models

2022-09-16 · Ben Prystawski, Paul Thibodeau, Christopher Potts, Noah D. Goodman

Probabilistic models of language understanding are valuable tools for investigating human language use. However, they need to be hand-designed for a particular domain. In contrast, large language models (LLMs) are traine…

JingFang: A Traditional Chinese Medicine Large Language Model of Expert-Level Medical Diagnosis and Syndrome Differentiation-Based Treatment

2025-02-04 · Yehan Yan, Tianhao Ma, Ruotai Li, Xinhan Zheng 외

Traditional Chinese medicine (TCM) plays a vital role in health protection and disease treatment, but its practical application requires extensive medical knowledge and clinical experience. Existing TCM Large Language Mo…

DiagnosticLanguage ModelingLanguage ModellingLarge Language Model+1

Probabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex Questions

2023-11-23 · Shulin Cao, Jiajie Zhang, Jiaxin Shi, Xin Lv 외

Large language models (LLMs) are capable of answering knowledge-intensive complex questions with chain-of-thought (CoT) reasoning. However, they tend to generate factually incorrect reasoning steps when the required know…

Retrieval

Enhancing Factual Accuracy and Citation Generation in LLMs via Multi-Stage Self-Verification

2025-09-06 · Fernando Gabriela García, Qiyang Shi, Zilin Feng arxiv

This research introduces VeriFact-CoT (Verified Factual Chain-of-Thought), a novel method designed to address the pervasive issues of hallucination and the absence of credible citation sources in Large Language Models (L…

Fact Verification