paper-with-me

Papers

Apollo: A Lightweight Multilingual Medical LLM towards Democratizing Medical AI to 6B People

2024-03-06 · Xidong Wang, Nuo Chen, Junyin Chen, Yidong Wang, Guorui Zhen, Chunxian Zhang, Xiangbo Wu, Yan Hu, Anningzhe Gao, Xiang Wan, Haizhou Li, Benyou Wang

Despite the vast repository of global medical knowledge predominantly being in English, local languages are crucial for delivering tailored healthcare services, particularly in areas with limited medical resources. To extend the reach of medical AI advancements to a broader population, we aim to develop medical LLMs across the six most widely spoken languages, encompassing a global population of 6.1 billion. This effort culminates in the creation of the ApolloCorpora multilingual medical dataset and the XMedBench benchmark. In the multilingual medical benchmark, the released Apollo models, at various relatively-small sizes (i.e., 0.5B, 1.8B, 2B, 6B, and 7B), achieve the best performance among models of equivalent size. Especially, Apollo-7B is the state-of-the-art multilingual medical LLMs up to 70B. Additionally, these lite models could be used to improve the multi-lingual medical capabilities of larger models without fine-tuning in a proxy-tuning fashion. We will open-source training corpora, code, model weights and evaluation benchmark.

📄 PDF Abstract BibTeX arXiv:2403.03640

Code (1)

freedomintelligence/apollo 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

Apollo Please enter a description about the method here

Similar Papers 제목 키워드 기반

Towards Democratizing Multilingual Large Language Models For Medicine Through A Two-Stage Instruction Fine-tuning Approach

2024-09-09 · Meng Zhou, Surajsinh Parmar, Anubhav Bhatti

Open-source, multilingual medical large language models (LLMs) have the potential to serve linguistically diverse populations across different regions. Adapting generic LLMs for healthcare often requires continual pretra…

Computational EfficiencyContinual PretrainingMultiple-choice

A multimodal and temporal foundation model for virtual patient representations at healthcare system scale

2026-04-20 · Andrew Zhang, Tong Ding, Sophia J. Wagner, Caiwei Tian 외 arxiv

Modern medicine generates vast multimodal data across siloed systems, yet no existing model integrates the full breadth and temporal depth of the clinical record into a unified patient representation. We introduce Apollo…

Semantic Similarity

Efficiently Democratizing Medical LLMs for 50 Languages via a Mixture of Language Family Experts

2024-10-14 · Guorui Zheng, Xidong Wang, Juhao Liang, Nuo Chen 외

Adapting medical Large Language Models to local languages can reduce barriers to accessing healthcare services, but data scarcity remains a significant challenge, particularly for low-resource languages. To address this,…

Mixture-of-Experts

U-RWKV: Lightweight medical image segmentation with direction-adaptive RWKV

2025-07-15 · Hongbo Ye, Fenghe Tang, Peiang Zhao, Zhen Huang 외

Achieving equity in healthcare accessibility requires lightweight yet high-performance solutions for medical image segmentation, particularly in resource-limited settings. Existing methods like U-Net and its variants oft…

Computational EfficiencyImage SegmentationLong-range modelingMedical Image Segmentation+1

BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains

2024-02-15 · Yanis Labrak, Adrien Bazoge, Emmanuel Morin, Pierre-Antoine Gourraud 외

Large Language Models (LLMs) have demonstrated remarkable versatility in recent years, offering potential applications across specialized domains such as healthcare and medicine. Despite the availability of various open-…

Few-Shot LearningMedical Question AnsweringQuantizationQuestion Answering+1