paper-with-me

홈 › Papers

Building a Silver-Standard Dataset from NICE Guidelines for Clinical LLMs

2025-11-02 · Qing Ding, Eric Hua Qing Zhang, Felix Jozsa, Julia Ive arxiv

Large language models (LLMs) are increasingly used in healthcare, yet standardised benchmarks for evaluating guideline-based clinical reasoning are missing. This study introduces a validated dataset derived from publicly available guidelines across multiple diagnoses. The dataset was created with the help of GPT and contains realistic patient scenarios, as well as clinical questions. We benchmark a range of recent popular LLMs to showcase the validity of our dataset. The framework supports systematic evaluation of LLMs' clinical utility and guideline adherence.

📄 PDF Abstract BibTeX arXiv:2511.01053

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

UD-English-CHILDES: A Collected Resource of Gold and Silver Universal Dependencies Trees for Child Language Interactions

2025-04-28 · Xiulin Yang, Zhuoxuan Ju, Lanni Bu, Zoey Liu 외

CHILDES is a widely used resource of transcribed child and child-directed speech. This paper introduces UD-English-CHILDES, the first officially released Universal Dependencies (UD) treebank derived from previously depen…

GRAIN-S: Manually Annotated Syntax for German Interviews

2020-05-01 · LREC 2020 5 · Agnieszka Falenska, Zolt{\'a}n Czesznak, Kerstin Jung, Moritz V{\"o}lkel 외

We present GRAIN-S, a set of manually created syntactic annotations for radio interviews in German. The dataset extends an existing corpus GRAIN and comes with constituency and dependency trees for six interviews. The ra…

Learning with Silver Standard Data for Zero-shot Relation Extraction

2022-11-25 · Tianyin Wang, Jianwei Wang, Ziqian Zeng

The superior performance of supervised relation extraction (RE) methods heavily relies on a large amount of gold standard data. Recent zero-shot relation extraction methods converted the RE task to other NLP tasks and us…

RelationRelation Extraction

Building a Synthetic Biomedical Research Article Citation Linkage Corpus

2022-06-01 · LREC 2022 6 · Sudipta Singha Roy, Robert E. Mercer

Citations are frequently used in publications to support the presented results and to demonstrate the previous discoveries while also assisting the reader in following the chronological progression of information through…

Semantic SimilaritySemantic Textual SimilaritySentenceSentence Embedding+1

Pixel-level Counterfactual Contrastive Learning for Medical Image Segmentation

2026-03-17 · Marceau Lafargue-Hauret, Raghav Mehta, Fabio De Sousa Ribeiro, Mélanie Roschewitz 외 arxiv

Image segmentation relies on large annotated datasets, which are expensive and slow to produce. Silver-standard (AI-generated) labels are easier to obtain, but they risk introducing bias. Self-supervised learning, needin…

Medical Image SegmentationSelf-Supervised LearningRepresentation LearningContrastive Learning