paper-with-me

홈 › Papers

MedThink: Enhancing Diagnostic Accuracy in Small Models via Teacher-Guided Reasoning Correction

2026-04-09 · Xinchun Su, Chunxu Luo, Lipeng Ma, Yixuan Li, Weidong Yang arxiv

Accurate clinical diagnosis requires extensive domain knowledge and complex clinical reasoning capabilities. Although large language models (LLMs) hold great potential for clinical reasoning, their high computational and memory requirements limit their deployment in resource-constrained environments. Knowledge distillation (KD) can compress LLM capabilities into smaller models, but traditional KD merely transfers superficial answer patterns and fails to preserve the structured reasoning required for reliable diagnosis. To address this, we propose a two-stage distillation framework, MedThink, designed to cultivate robust clinical reasoning in small language models (SLMs). In the first stage, a teacher LLM screens data and injects domain-knowledge explanations to fine-tune a student model, establishing a knowledge foundation. In the second stage, the teacher evaluates the student's errors, generates reasoning chains linking knowledge to correct answers, and refines the student's diagnostic reasoning through a second round of fine-tuning. We evaluate MedThink on general medical benchmarks and a gastroenterology dataset comprising 955 question-answer pairs. Experiments demonstrate that MedThink outperforms six distillation strategies in all benchmarks: achieving an improvement of up to 12.7% over the student baseline in general tasks, and reaching a total top accuracy of 56.4% in gastroenterology evaluation. This indicates that iterative distillation centered on reasoning can significantly enhance the diagnostic accuracy and generalization capabilities of SLMs whilst maintaining computational efficiency. Our code and data are publicly available at https://github.com/destinybird/PrecisionBoost.

📄 PDF Abstract BibTeX arXiv:2605.08094

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyKnowledge Distillation

Similar Papers 제목 키워드 기반

Automating Expert-Level Medical Reasoning Evaluation of Large Language Models

2025-07-10 · Shuang Zhou, Wenya Xie, Jiaxi Li, Zaifu Zhan 외 arxiv

As large language models (LLMs) become increasingly integrated into clinical decision-making, ensuring transparent and trustworthy reasoning is essential. However, existing evaluation strategies of LLMs' medical reasonin…

MedThink: Explaining Medical Visual Question Answering via Multimodal Decision-Making Rationale

2024-04-18 · Xiaotang Gai, Chenyi Zhou, Jiaxiang Liu, Yang Feng 외

Medical Visual Question Answering (MedVQA), which offers language responses to image-based medical inquiries, represents a challenging task and significant advancement in healthcare. It assists medical experts to swiftly…

Decision MakingMedical Visual Question AnsweringQuestion AnsweringVisual Question Answering+1

Evaluating Knowledge Transfer in Neural Network for Medical Images

2020-08-31 · Sina Akbarian, Laleh Seyyed-Kalantari, Farzad Khalvati, Elham Dolatabadi

Deep learning and knowledge transfer techniques have permeated the field of medical imaging and are considered as key approaches for revolutionizing diagnostic imaging practices. However, there are still challenges for t…

DiagnosticTransfer Learning

DiaCDM: Cognitive Diagnosis in Teacher-Student Dialogues using the Initiation-Response-Evaluation Framework

2025-09-29 · Rui Jia, Yuang Wei, Ruijia Li, Yuan-Hao Jiang 외 arxiv

While cognitive diagnosis (CD) effectively assesses students' knowledge mastery from structured test data, applying it to real-world teacher-student dialogues presents two fundamental challenges. Traditional CD models la…

Adversarial Distillation Based on Slack Matching and Attribution Region Alignment

2024-01-01 · CVPR 2024 1 · Shenglin Yin, Zhen Xiao, Mingxuan Song, Jieyi Long

Adversarial distillation (AD) is a highly effective method for enhancing the robustness of small models. Contrary to expectations a high-performing teacher model does not always result in a more robust student model.…