paper-with-me

홈 › Papers

Clinical Language Understanding Evaluation (CLUE)

2022-09-28 · Travis R. Goodwin, Dina Demner-Fushman

Clinical language processing has received a lot of attention in recent years, resulting in new models or methods for disease phenotyping, mortality prediction, and other tasks. Unfortunately, many of these approaches are tested under different experimental settings (e.g., data sources, training and testing splits, metrics, evaluation criteria, etc.) making it difficult to compare approaches and determine state-of-the-art. To address these issues and facilitate reproducibility and comparison, we present the Clinical Language Understanding Evaluation (CLUE) benchmark with a set of four clinical language understanding tasks, standard training, development, validation and testing sets derived from MIMIC data, as well as a software toolkit. It is our hope that these data will enable direct comparison between approaches, improve reproducibility, and reduce the barrier-to-entry for developing novel models or methods for these clinical language understanding tasks.

📄 PDF Abstract BibTeX arXiv:2209.14377

Code (0)

등록된 구현이 없습니다.

Tasks

Mortality Prediction

Similar Papers 제목 키워드 기반

CLUE: A Chinese Language Understanding Evaluation Benchmark

2020-04-13 · COLING 2020 8 · Liang Xu, Hai Hu, Xuanwei Zhang, Lu Li 외

The advent of natural language understanding (NLU) benchmarks for English, such as GLUE and SuperGLUE allows new NLU models to be evaluated across a diverse set of tasks. These comprehensive benchmarks have facilitated a…

General ClassificationMachine Reading ComprehensionNatural Language UnderstandingReading Comprehension+3

CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding

2024-12-16 · Guo Chen, Yicheng Liu, Yifei HUANG, Yuping He 외

Most existing video understanding benchmarks for multimodal large language models (MLLMs) focus only on short videos. The limited number of benchmarks for long video understanding often rely solely on multiple-choice que…

HallucinationMultiple-choiceQuestion AnsweringVideo Understanding

CLUES: Few-Shot Learning Evaluation in Natural Language Understanding

2021-11-04 · Subhabrata Mukherjee, Xiaodong Liu, Guoqing Zheng, Saghar Hosseini 외

Most recent progress in natural language understanding (NLU) has been driven, in part, by benchmarks such as GLUE, SuperGLUE, SQuAD, etc. In fact, many NLU models have now matched or exceeded "human-level" performance on…

Few-Shot LearningNatural Language Understanding

Can Large Language Model Comprehend Ancient Chinese? A Preliminary Test on ACLUE

2023-10-14 · Yixuan Zhang, Haonan Li

Large language models (LLMs) have showcased remarkable capabilities in understanding and generating language. However, their ability in comprehending ancient languages, particularly ancient Chinese, remains largely unexp…

Language ModelingLanguage ModellingLarge Language Model

Italian Crossword Generator: Enhancing Education through Interactive Word Puzzles

2023-11-27 · Kamyar Zeinalipour, Tommaso laquinta, Asya Zanollo, Giovanni Angelini 외

Educational crosswords offer numerous benefits for students, including increased engagement, improved understanding, critical thinking, and memory retention. Creating high-quality educational crosswords can be challengin…

Few-Shot LearningZero-Shot Learning