paper-with-me

홈 › Papers

From Generalist to Specialist: Improving Large Language Models for Medical Physics Using ARCoT

2024-05-17 · Jace Grandinetti, Rafe McBeth

Large Language Models (LLMs) have achieved remarkable progress, yet their application in specialized fields, such as medical physics, remains challenging due to the need for domain-specific knowledge. This study introduces ARCoT (Adaptable Retrieval-based Chain of Thought), a framework designed to enhance the domain-specific accuracy of LLMs without requiring fine-tuning or extensive retraining. ARCoT integrates a retrieval mechanism to access relevant domain-specific information and employs step-back and chain-of-thought prompting techniques to guide the LLM's reasoning process, ensuring more accurate and context-aware responses. Benchmarking on a medical physics multiple-choice exam, our model outperformed standard LLMs and reported average human performance, demonstrating improvements of up to 68% and achieving a high score of 90%. This method reduces hallucinations and increases domain-specific performance. The versatility and model-agnostic nature of ARCoT make it easily adaptable to various domains, showcasing its significant potential for enhancing the accuracy and reliability of LLMs in specialized fields.

📄 PDF Abstract BibTeX arXiv:2405.11040

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingMultiple-choiceRetrieval

Similar Papers 제목 키워드 기반

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

2026-05-28 · Yanan Wang, Shuaicong Hu, Jian Liu, Guohui Zhou 외 arxiv

The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-specific medical specialist models become obsolete? We argue that the fut…

Super-Generalist: Towards Comprehensive and Accurate Medical Image Understanding via Generalist-Specialist Synergy

2026-07-10 · Shaoteng Zhang, Weiwei Cao, Wanxing Chang, Yutong Xie 외 arxiv

Medical images require comprehensive and accurate interpretation to support the diagnosis of diverse clincial conditions. Recent vision-language generalist models offer broad task coverage and promising zero-shot capabil…

From Specialist to Generalist: Unlocking SAM's Learning Potential on Unlabeled Medical Images

2026-01-25 · Vi Vu, Thanh-Huy Nguyen, Tien-Thinh Nguyen, Ba-Thinh Lam 외 arxiv

Foundation models like the Segment Anything Model (SAM) show strong generalization, yet adapting them to medical images remains difficult due to domain shift, scarce labels, and the inability of Parameter-Efficient Fine-…

parameter-efficient fine-tuningMedical Image SegmentationPolyp Segmentation

CLARIFY: A Specialist-Generalist Framework for Accurate and Lightweight Dermatological Visual Question Answering

2025-08-25 · Aranya Saha, Tanvir Ahmed Khan, Ismam Nur Swapnil, Mohammad Ariful Haque arxiv

Vision-language models (VLMs) have shown significant potential for medical tasks; however, their general-purpose nature can limit specialized diagnostic accuracy, and their large size poses substantial inference costs fo…

Visual Question AnsweringComputational Efficiency

TAGS: A Test-Time Generalist-Specialist Framework with Retrieval-Augmented Reasoning and Verification

2025-05-23 · Jianghao Wu, Feilong Tang, Yulong Li, Ming Hu 외

Recent advances such as Chain-of-Thought prompting have significantly improved large language models (LLMs) in zero-shot medical reasoning. However, prompting-based methods often remain shallow and unstable, while fine-t…

MedQA