paper-with-me

홈 › Papers

Hippocrates: An Open-Source Framework for Advancing Large Language Models in Healthcare

2024-04-25 · Emre Can Acikgoz, Osman Batur İnce, Rayene Bench, Arda Anıl Boz, İlker Kesen, Aykut Erdem, Erkut Erdem

The integration of Large Language Models (LLMs) into healthcare promises to transform medical diagnostics, research, and patient care. Yet, the progression of medical LLMs faces obstacles such as complex training requirements, rigorous evaluation demands, and the dominance of proprietary models that restrict academic exploration. Transparent, comprehensive access to LLM resources is essential for advancing the field, fostering reproducibility, and encouraging innovation in healthcare AI. We present Hippocrates, an open-source LLM framework specifically developed for the medical domain. In stark contrast to previous efforts, it offers unrestricted access to its training datasets, codebase, checkpoints, and evaluation protocols. This open approach is designed to stimulate collaborative research, allowing the community to build upon, refine, and rigorously evaluate medical LLMs within a transparent ecosystem. Also, we introduce Hippo, a family of 7B models tailored for the medical domain, fine-tuned from Mistral and LLaMA2 through continual pre-training, instruction tuning, and reinforcement learning from human and AI feedback. Our models outperform existing open medical LLMs models by a large-margin, even surpassing models with 70B parameters. Through Hippocrates, we aspire to unlock the full potential of LLMs not just to advance medical knowledge and patient care but also to democratize the benefits of AI research in healthcare, making them available across the globe.

📄 PDF Abstract BibTeX arXiv:2404.16621

Code (1)

hiyouga/llama-factory 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Sufficient and necessary causation are dual

2017-10-25 · Robert Künnemann

Causation has been the issue of philosophic debate since Hippocrates. Recent work defines actual causation in terms of Pearl/Halpern's causality framework, formalizing necessary causes (IJCAI'15). This has inspired causa…

Relation

The Open Source Advantage in Large Language Models (LLMs)

2024-12-16 · Jiya Manchanda, Laura Boettcher, Matheus Westphalen, Jasser Jasser

Large language models (LLMs) have rapidly advanced natural language processing, driving significant breakthroughs in tasks such as text generation, machine translation, and domain-specific reasoning. The field now faces …

Computational EfficiencyMachine TranslationPositionText Generation

A Review of DeepSeek Models' Key Innovative Techniques

2025-03-14 · Chengen Wang, Murat Kantarcioglu

DeepSeek-V3 and DeepSeek-R1 are leading open-source Large Language Models (LLMs) for general-purpose tasks and reasoning, achieving performance comparable to state-of-the-art closed-source models from companies like Open…

Mixture-of-Expertsreinforcement-learningReinforcement Learning

O-Researcher: An Open Ended Deep Research Model via Multi-Agent Distillation and Agentic RL

2026-01-07 · Yi Yao, He Zhu, Piaohong Wang, Jincheng Ren 외 arxiv

The performance gap between closed-source and open-source large language models (LLMs) is largely attributed to disparities in access to high-quality training data. To bridge this gap, we introduce a novel framework for …

Reinforcement Learning

WanJuanSiLu: A High-Quality Open-Source Webtext Dataset for Low-Resource Languages

2025-01-24 · JIA YU, Fei Yuan, Rui Min, Jing Yu 외

This paper introduces the open-source dataset WanJuanSiLu, designed to provide high-quality training corpora for low-resource languages, thereby advancing the research and development of multilingual models. To achieve t…

Diversity