paper-with-me

홈 › Papers

ChiMed-GPT: A Chinese Medical Large Language Model with Full Training Regime and Better Alignment to Human Preferences

2023-11-10 · Yuanhe Tian, Ruyi Gan, Yan Song, Jiaxing Zhang, Yongdong Zhang

Recently, the increasing demand for superior medical services has highlighted the discrepancies in the medical infrastructure. With big data, especially texts, forming the foundation of medical services, there is an exigent need for effective natural language processing (NLP) solutions tailored to the healthcare domain. Conventional approaches leveraging pre-trained models present promising results in this domain and current large language models (LLMs) offer advanced foundation for medical text processing. However, most medical LLMs are trained only with supervised fine-tuning (SFT), even though it efficiently empowers LLMs to understand and respond to medical instructions but is ineffective in learning domain knowledge and aligning with human preference. In this work, we propose ChiMed-GPT, a new benchmark LLM designed explicitly for Chinese medical domain, and undergoes a comprehensive training regime with pre-training, SFT, and RLHF. Evaluations on tasks including information extraction, question answering, and dialogue generation demonstrate ChiMed-GPT's superior performance over general domain LLMs. Furthermore, we analyze possible biases through prompting ChiMed-GPT to perform attitude scales regarding discrimination of patients, so as to contribute to further responsible development of LLMs in the medical domain. The code and model are released at https://github.com/synlp/ChiMed-GPT.

📄 PDF Abstract BibTeX arXiv:2311.06025

Code (1)

synlp/chimed-gpt 공식 구현 pytorch

Tasks

Dialogue GenerationLanguage ModelingLanguage ModellingLarge Language ModelQuestion Answering

Methods 이 논문이 사용한 방법론

SFT Shrink and Fine-Tune, or SFT, is a type of distillation that avoids explicit distillation by copying parameters to a student student model and then fine-tuning.…

Similar Papers 제목 키워드 기반

ChiMed 2.0: Advancing Chinese Medical Dataset in Facilitating Large Language Modeling

2025-07-21 · Yuanhe Tian, Junjie Liu, Zhizhou Kou, Yuxiang Li 외 arxiv

Building high-quality data resources is crucial for advancing artificial intelligence research and applications in specific domains, particularly in the Chinese medical domain. Existing Chinese medical datasets are limit…

Reinforcement Learning

ChiMed: A Chinese Medical Corpus for Question Answering

2019-08-01 · WS 2019 8 · Yuanhe Tian, Weicheng Ma, Fei Xia, Yan Song

Question answering (QA) is a challenging task in natural language processing (NLP), especially when it is applied to specific domains. While models trained in the general domain can be adapted to a new target domain, the…

Question Answering

Qilin-Med-VL: Towards Chinese Large Vision-Language Model for General Healthcare

2023-10-27 · Junling Liu, ZiMing Wang, Qichen Ye, Dading Chong 외

Large Language Models (LLMs) have introduced a new era of proficiency in comprehending complex healthcare and biomedical topics. However, there is a noticeable lack of models in languages other than English and models th…

Language ModelingLanguage Modelling

Qilin-Med: Multi-stage Knowledge Injection Advanced Medical Large Language Model

2023-10-13 · Qichen Ye, Junling Liu, Dading Chong, Peilin Zhou 외

Integrating large language models (LLMs) into healthcare holds great potential but faces challenges. Pre-training LLMs from scratch for domains like medicine is resource-heavy and often unfeasible. On the other hand, sol…

Knowledge GraphsLanguage ModelingLanguage ModellingLarge Language Model+4

MedChatZH: a Better Medical Adviser Learns from Better Instructions

2023-09-03 · Yang Tan, Mingchen Li, Zijie Huang, Huiqun Yu 외

Generative large language models (LLMs) have shown great success in various applications, including question-answering (QA) and dialogue systems. However, in specialized domains like traditional Chinese medical QA, these…

Question Answering