paper-with-me

홈 › Papers

Developing Healthcare Language Model Embedding Spaces

2024-03-28 · Niall Taylor, Dan Schofield, Andrey Kormilitzin, Dan W Joyce, Alejo Nevado-Holgado

Pre-trained Large Language Models (LLMs) often struggle on out-of-domain datasets like healthcare focused text. We explore specialized pre-training to adapt smaller LLMs to different healthcare datasets. Three methods are assessed: traditional masked language modeling, Deep Contrastive Learning for Unsupervised Textual Representations (DeCLUTR), and a novel pre-training objective utilizing metadata categories from the healthcare settings. These schemes are evaluated on downstream document classification tasks for each dataset, with additional analysis of the resultant embedding spaces. Contrastively trained models outperform other approaches on the classification tasks, delivering strong performance from limited labeled data and with fewer model parameter updates required. While metadata-based pre-training does not further improve classifications across the datasets, it yields interesting embedding cluster separability. All domain adapted LLMs outperform their publicly available general base LLM, validating the importance of domain-specialization. This research illustrates efficient approaches to instill healthcare competency in compact LLMs even under tight computational budgets, an essential capability for responsible and sustainable deployment in local healthcare settings. We provide pre-training guidelines for specialized healthcare LLMs, motivate continued inquiry into contrastive objectives, and demonstrates adaptation techniques to align small LLMs with privacy-sensitive medical tasks.

📄 PDF Abstract BibTeX arXiv:2403.19802

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningDocument ClassificationLanguage ModelingLanguage ModellingMasked Language Modelingmodel

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…
BASE 설명 없음

Similar Papers 제목 키워드 기반

Unsupervised Cross-Modal Alignment of Speech and Text Embedding Spaces

2018-05-18 · NeurIPS 2018 12 · Yu-An Chung, Wei-Hung Weng, Schrasing Tong, James Glass

Recent research has shown that word embedding spaces learned from text corpora of different languages can be aligned without any parallel data supervision. Inspired by the success in unsupervised cross-lingual word embed…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Cross-Lingual Word Embeddingscross-modal alignment+6

Snomed2Vec: Random Walk and Poincaré Embeddings of a Clinical Knowledge Base for Healthcare Analytics

2019-07-19 · Khushbu Agarwal, Tome Eftimov, Raghavendra Addanki, Sutanay Choudhury 외

Representation learning methods that transform encoded data (e.g., diagnosis and drug codes) into continuous vector spaces (i.e., vector embeddings) are critical for the application of deep learning in healthcare. Initia…

Clinical KnowledgeLink PredictionNode ClassificationRepresentation Learning

A Comparative Study on Structural and Semantic Properties of Sentence Embeddings

2020-09-23 · Alexander Kalinowski, Yuan An

Sentence embeddings encode natural language sentences as low-dimensional dense vectors. A great deal of effort has been put into using sentence embeddings to improve several important natural language processing tasks. R…

Knowledge GraphsRelationRelation ExtractionSentence+3

AURORA: Contextual Orthogonalization for Geometric Representation Learning in Healthcare Foundation Models

2026-05-18 · Yuanyun Zhang, Shi Li arxiv

Recent healthcare foundation models have achieved strong predictive performance through large scale self supervised learning, yet their latent representations frequently entangle physiologic severity, intervention intens…

Representation Learning

Learning in Hilbert vs. Banach Spaces: A Measure Embedding Viewpoint

2011-12-01 · NeurIPS 2011 12 · Kenji Fukumizu, Gert R. Lanckriet, Bharath K. Sriperumbudur

The goal of this paper is to investigate the advantages and disadvantages of learning in Banach spaces over Hilbert spaces. While many works have been carried out in generalizing Hilbert methods to Banach spaces, in this…